Official AAPA release: processed training data and A-GRPO checkpoints for adversarially anchored preference alignment.
Jingleqian
Jingleqian
AI & ML interests
None yet
Recent Activity
new activity about 1 month ago
Jingleqian/AAPA-data:Add dataset card and link to paper/code new activity about 1 month ago
Jingleqian/AAPA-8B:Add pipeline tag and link to paper new activity about 1 month ago
Jingleqian/AAPA-06B:Update model card with paper link and pipeline tagOrganizations
None yet