2026-01-263 saved articles | Back | Library |
|---|
| Article | Read |
|---|---|
| Recursive Self-Aggregation: Deep Thinking and Test-Time Scaling for LLM Reasoning RSA aggregates multiple chains-of-thought across iterations to unlock deep reasoning in LLMs. Aggregation-aware RL further boosts performance across math, coding, and knowledge tasks. | |
| rsa-llm.github.io | |
| GitHub - tobi/qmd: mini cli search engine for your docs, knowledge bases, meeting notes, whatever. Tracking current sota approaches while being all local mini cli search engine for your docs, knowledge bases, meeting notes, whatever. Tracking current sota approaches while being all local - tobi/qmd |