-
2.5 Years in Class: A Multimodal Textbook for Vision-Language Pretraining
Paper • 2501.00958 • Published • 99 -
OS-Genesis: Automating GUI Agent Trajectory Construction via Reverse Task Synthesis
Paper • 2412.19723 • Published • 82 -
PERSE: Personalized 3D Generative Avatars from A Single Portrait
Paper • 2412.21206 • Published • 17 -
Training Software Engineering Agents and Verifiers with SWE-Gym
Paper • 2412.21139 • Published • 21
Coning
Nycbro8
AI & ML interests
None yet
Recent Activity
updated
a collection
about 1 month ago
Fvb
upvoted
a
paper
about 1 month ago
MentalLLaMA: Interpretable Mental Health Analysis on Social Media with
Large Language Models
upvoted
a
collection
about 1 month ago
IBE
Organizations
None yet
Collections
1
models
None public yet
datasets
None public yet