Collections
Discover the best community collections!
Collections including paper arxiv:2608.02287
-
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources
Paper • 2606.29538 • Published • 144 -
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation
Paper • 2608.02287 • Published • 31 -
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
Paper • 2608.05139 • Published • 27 -
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure
Paper • 2608.11079 • Published • 17
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 1.82k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89
-
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks
Paper • 2608.01964 • Published • 181 -
DAPD: Dual-Anchored Policy Distillation
Paper • 2608.01735 • Published • 151 -
Progressive Agent Skill Generation via Reinforcement Learning
Paper • 2608.01678 • Published • 60 -
UEmbed: Unified Sparse and Dense Multimodal Embeddings
Paper • 2608.02583 • Published • 50
-
Skywork-SWE: Unveiling Data Scaling Laws for Software Engineering in LLMs
Paper • 2506.19290 • Published • 53 -
Data Efficacy for Language Model Training
Paper • 2506.21545 • Published • 11 -
Easy Dataset: A Unified and Extensible Framework for Synthesizing LLM Fine-Tuning Data from Unstructured Documents
Paper • 2507.04009 • Published • 55 -
RefineX: Learning to Refine Pre-training Data at Scale from Expert-Guided Programs
Paper • 2507.03253 • Published • 19
-
LongHorizon-Harness: Advancing Long-Horizon Agents for Real-World Tasks
Paper • 2608.01964 • Published • 181 -
DAPD: Dual-Anchored Policy Distillation
Paper • 2608.01735 • Published • 151 -
Progressive Agent Skill Generation via Reinforcement Learning
Paper • 2608.01678 • Published • 60 -
UEmbed: Unified Sparse and Dense Multimodal Embeddings
Paper • 2608.02583 • Published • 50
-
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources
Paper • 2606.29538 • Published • 144 -
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation
Paper • 2608.02287 • Published • 31 -
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning
Paper • 2608.05139 • Published • 27 -
SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure
Paper • 2608.11079 • Published • 17
-
Skywork-SWE: Unveiling Data Scaling Laws for Software Engineering in LLMs
Paper • 2506.19290 • Published • 53 -
Data Efficacy for Language Model Training
Paper • 2506.21545 • Published • 11 -
Easy Dataset: A Unified and Extensible Framework for Synthesizing LLM Fine-Tuning Data from Unstructured Documents
Paper • 2507.04009 • Published • 55 -
RefineX: Learning to Refine Pre-training Data at Scale from Expert-Guided Programs
Paper • 2507.03253 • Published • 19
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 1.82k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89