-
Attention Is All You Need
Paper • 1706.03762 • Published • 134 -
Scaling Laws for Neural Language Models
Paper • 2001.08361 • Published • 10 -
Training Compute-Optimal Large Language Models
Paper • 2203.15556 • Published • 12 -
Analogy Generation by Prompting Large Language Models: A Case Study of InstructGPT
Paper • 2210.04186 • Published
Collections
Discover the best community collections!
Collections including paper arxiv:2506.02153
-
A Survey of Small Language Models
Paper • 2410.20011 • Published • 46 -
Small Language Models are the Future of Agentic AI
Paper • 2506.02153 • Published • 25 -
AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System
Paper • 2402.15538 • Published • 6 -
Small Language Models: Survey, Measurements, and Insights
Paper • 2409.15790 • Published • 1
-
LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries
Paper • 2508.15760 • Published • 47 -
LiveMCPBench: Can Agents Navigate an Ocean of MCP Tools?
Paper • 2508.01780 • Published • 21 -
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs
Paper • 2304.08244 • Published • 1 -
AgentFly: Fine-tuning LLM Agents without Fine-tuning LLMs
Paper • 2508.16153 • Published • 162
-
The Leaderboard Illusion
Paper • 2504.20879 • Published • 71 -
SmolVLM: Redefining small and efficient multimodal models
Paper • 2504.05299 • Published • 210 -
Seedance 1.0: Exploring the Boundaries of Video Generation Models
Paper • 2506.09113 • Published • 109 -
Small Language Models are the Future of Agentic AI
Paper • 2506.02153 • Published • 25
-
Neural Machine Translation by Jointly Learning to Align and Translate
Paper • 1409.0473 • Published • 7 -
Attention Is All You Need
Paper • 1706.03762 • Published • 134 -
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Paper • 1810.04805 • Published • 32 -
Hierarchical Reasoning Model
Paper • 2506.21734 • Published • 54
-
Why Language Models Hallucinate
Paper • 2509.04664 • Published • 200 -
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
Paper • 2508.21184 • Published • 3 -
Reflect, Retry, Reward: Self-Improving LLMs via Reinforcement Learning
Paper • 2505.24726 • Published • 283 -
Small Language Models are the Future of Agentic AI
Paper • 2506.02153 • Published • 25
-
Attention Is All You Need
Paper • 1706.03762 • Published • 134 -
Scaling Laws for Neural Language Models
Paper • 2001.08361 • Published • 10 -
Training Compute-Optimal Large Language Models
Paper • 2203.15556 • Published • 12 -
Analogy Generation by Prompting Large Language Models: A Case Study of InstructGPT
Paper • 2210.04186 • Published
-
Neural Machine Translation by Jointly Learning to Align and Translate
Paper • 1409.0473 • Published • 7 -
Attention Is All You Need
Paper • 1706.03762 • Published • 134 -
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Paper • 1810.04805 • Published • 32 -
Hierarchical Reasoning Model
Paper • 2506.21734 • Published • 54
-
A Survey of Small Language Models
Paper • 2410.20011 • Published • 46 -
Small Language Models are the Future of Agentic AI
Paper • 2506.02153 • Published • 25 -
AgentLite: A Lightweight Library for Building and Advancing Task-Oriented LLM Agent System
Paper • 2402.15538 • Published • 6 -
Small Language Models: Survey, Measurements, and Insights
Paper • 2409.15790 • Published • 1
-
Why Language Models Hallucinate
Paper • 2509.04664 • Published • 200 -
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
Paper • 2508.21184 • Published • 3 -
Reflect, Retry, Reward: Self-Improving LLMs via Reinforcement Learning
Paper • 2505.24726 • Published • 283 -
Small Language Models are the Future of Agentic AI
Paper • 2506.02153 • Published • 25
-
LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries
Paper • 2508.15760 • Published • 47 -
LiveMCPBench: Can Agents Navigate an Ocean of MCP Tools?
Paper • 2508.01780 • Published • 21 -
API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs
Paper • 2304.08244 • Published • 1 -
AgentFly: Fine-tuning LLM Agents without Fine-tuning LLMs
Paper • 2508.16153 • Published • 162
-
The Leaderboard Illusion
Paper • 2504.20879 • Published • 71 -
SmolVLM: Redefining small and efficient multimodal models
Paper • 2504.05299 • Published • 210 -
Seedance 1.0: Exploring the Boundaries of Video Generation Models
Paper • 2506.09113 • Published • 109 -
Small Language Models are the Future of Agentic AI
Paper • 2506.02153 • Published • 25