-
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Paper • 2402.04252 • Published • 26 -
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
Paper • 2402.03749 • Published • 13 -
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Paper • 2402.04615 • Published • 43 -
EfficientViT-SAM: Accelerated Segment Anything Model Without Performance Loss
Paper • 2402.05008 • Published • 22
Collections
Discover the best community collections!
Collections including paper arxiv:2403.19270
-
LoRA+: Efficient Low Rank Adaptation of Large Models
Paper • 2402.12354 • Published • 6 -
The FinBen: An Holistic Financial Benchmark for Large Language Models
Paper • 2402.12659 • Published • 21 -
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization
Paper • 2402.13249 • Published • 13 -
TrustLLM: Trustworthiness in Large Language Models
Paper • 2401.05561 • Published • 69
-
LLMLingua-2: Data Distillation for Efficient and Faithful Task-Agnostic Prompt Compression
Paper • 2403.12968 • Published • 25 -
PERL: Parameter Efficient Reinforcement Learning from Human Feedback
Paper • 2403.10704 • Published • 58 -
Alignment Studio: Aligning Large Language Models to Particular Contextual Regulations
Paper • 2403.09704 • Published • 32 -
RAFT: Adapting Language Model to Domain Specific RAG
Paper • 2403.10131 • Published • 71
-
sDPO: Don't Use Your Data All at Once
Paper • 2403.19270 • Published • 41 -
Advancing LLM Reasoning Generalists with Preference Trees
Paper • 2404.02078 • Published • 45 -
Learn Your Reference Model for Real Good Alignment
Paper • 2404.09656 • Published • 85 -
mDPO: Conditional Preference Optimization for Multimodal Large Language Models
Paper • 2406.11839 • Published • 38
-
Jamba: A Hybrid Transformer-Mamba Language Model
Paper • 2403.19887 • Published • 109 -
sDPO: Don't Use Your Data All at Once
Paper • 2403.19270 • Published • 41 -
ViTAR: Vision Transformer with Any Resolution
Paper • 2403.18361 • Published • 55 -
Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models
Paper • 2403.18814 • Published • 47
-
On the Societal Impact of Open Foundation Models
Paper • 2403.07918 • Published • 17 -
sDPO: Don't Use Your Data All at Once
Paper • 2403.19270 • Published • 41 -
Hallucinations or Attention Misdirection? The Path to Strategic Value Extraction in Business Using Large Language Models
Paper • 2402.14002 • Published -
Evaluating the Social Impact of Generative AI Systems in Systems and Society
Paper • 2306.05949 • Published • 9
-
InternLM2 Technical Report
Paper • 2403.17297 • Published • 32 -
sDPO: Don't Use Your Data All at Once
Paper • 2403.19270 • Published • 41 -
Learn Your Reference Model for Real Good Alignment
Paper • 2404.09656 • Published • 85 -
OpenBezoar: Small, Cost-Effective and Open Models Trained on Mixes of Instruction Data
Paper • 2404.12195 • Published • 12