-
RLHF Workflow: From Reward Modeling to Online RLHF
Paper β’ 2405.07863 β’ Published β’ 66 -
Chameleon: Mixed-Modal Early-Fusion Foundation Models
Paper β’ 2405.09818 β’ Published β’ 127 -
Meteor: Mamba-based Traversal of Rationale for Large Language and Vision Models
Paper β’ 2405.15574 β’ Published β’ 53 -
An Introduction to Vision-Language Modeling
Paper β’ 2405.17247 β’ Published β’ 87
Collections
Discover the best community collections!
Collections including paper arxiv:2407.16741
-
9.02kπ©βπ¨
AI Comic Factory
Create your own AI comic with a single prompt
-
762π¨βπ€
Face to All
AI filter for your portraits
-
flowfree/crypto-news-headlines
Viewer β’ Updated β’ 1.04k β’ 119 β’ 8 -
OpenDevin: An Open Platform for AI Software Developers as Generalist Agents
Paper β’ 2407.16741 β’ Published β’ 69
-
Attention Is All You Need
Paper β’ 1706.03762 β’ Published β’ 50 -
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Paper β’ 1810.04805 β’ Published β’ 16 -
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Paper β’ 1910.01108 β’ Published β’ 14 -
Language Models are Few-Shot Learners
Paper β’ 2005.14165 β’ Published β’ 12
-
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
Paper β’ 2403.03507 β’ Published β’ 183 -
Yi: Open Foundation Models by 01.AI
Paper β’ 2403.04652 β’ Published β’ 62 -
RLHF Can Speak Many Languages: Unlocking Multilingual Preference Optimization for LLMs
Paper β’ 2407.02552 β’ Published β’ 4 -
OpenDevin: An Open Platform for AI Software Developers as Generalist Agents
Paper β’ 2407.16741 β’ Published β’ 69
-
MegaScale: Scaling Large Language Model Training to More Than 10,000 GPUs
Paper β’ 2402.15627 β’ Published β’ 34 -
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
Paper β’ 2402.16822 β’ Published β’ 15 -
FuseChat: Knowledge Fusion of Chat Models
Paper β’ 2402.16107 β’ Published β’ 36 -
Multi-LoRA Composition for Image Generation
Paper β’ 2402.16843 β’ Published β’ 28
-
AgentOhana: Design Unified Data and Training Pipeline for Effective Agent Learning
Paper β’ 2402.15506 β’ Published β’ 14 -
AutoWebGLM: Bootstrap And Reinforce A Large Language Model-based Web Navigating Agent
Paper β’ 2404.03648 β’ Published β’ 24 -
Similarity is Not All You Need: Endowing Retrieval Augmented Generation with Multi Layered Thoughts
Paper β’ 2405.19893 β’ Published β’ 31 -
Parrot: Efficient Serving of LLM-based Applications with Semantic Variable
Paper β’ 2405.19888 β’ Published β’ 7