-
Chain-of-Thought Reasoning Without Prompting
Paper • 2402.10200 • Published • 111 -
Self-Discover: Large Language Models Self-Compose Reasoning Structures
Paper • 2402.03620 • Published • 117 -
Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences
Paper • 2404.03715 • Published • 62 -
Do language models plan ahead for future tokens?
Paper • 2404.00859 • Published • 3
Thomas Renkert
trenkert
AI & ML interests
None yet
Recent Activity
new activity about 5 hours ago
Atomic-Germ/Ornith-1.0-35B-A3B-NPU2:Update to FLM 1.0.2? liked a model 7 days ago
zai-org/GLM-5.3 liked a model about 1 month ago
moonshotai/Kimi-K3