AI Radar
Support
LiveUpdated 2026-09-19 18:08 UTC

arXiv.org

arXiv.org — Language models can produce plausible short proofs, but may still be unreliable on…

arXiv.org is Language models can produce plausible short proofs, but may still be unreliable on long-horizon research problems, where progress depends on a sequence of uncertain and interdependent decisions. We introduce Stellar…. It is ranked #481 on the AI Radar, in AI research, first seen 3 days ago and shared in 1 post (20.1k views).

Visit arxiv.org

What people said about arXiv.org on X

Banger paper from Google on agent harnesses for long-horizon tasks. (bookmark it) Google Research built a many-agent harness for long mathematical proofs, and it produced new results on open problems from FOCS and JMLR papers. Stellar Colosseum works in stages. It explores several proof strategies, waits for a…

@omarsar0, 3 days ago · 298 likes · see the post

Alternatives to arXiv.org

arXiv.org in numbers

FAQ

What is arXiv.org?

Language models can produce plausible short proofs, but may still be unreliable on long-horizon research problems, where progress depends on a sequence of uncertain and interdependent decisions. We introduce Stellar… It was first shared on X 3 days ago and is ranked #481 on the AI Radar.

Is arXiv.org free?

Yes, it is free to use.

Who shared arXiv.org?

1 account on X, including @omarsar0, in 1 post totalling 20.1k views.

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 27.2k posts from 4.8k X accounts over the last 14 days, 1.2k tools, 19 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-19 18:08 UTC. Full method.