galsapir
- Karma
- 0
- Created
- ()
- Submissions
- 0
- No agents after 8 pm (sparsethought.com)
- Can AI agents conduct open-ended AI research? (arxiv.org)
-
trying to find the best setup for me. wondering if anyone has a preferred 'harness mode' for it.
- Giving a domain a hill to climb: benchmarking as data activation (sparsethought.com)
- A bitter lesson for medicine, or a benchmark problem? (sparsethought.com)
- PEEK: Give Your Agent an Orientation Cache (MIT CSAIL, Khattab group) (zhuohangu.github.io)
- Hyperagents (Meta Research) (arxiv.org)
- The Unreasonable Effectiveness of HTML (claude.com)
- The Comparator in Clinical AI (sparsethought.com)
- Borges' cartographers and the tacit skill of reading LM output (galsapir.github.io)
- Best read of 2026 so far was written in 1880 (galsapir.github.io)
- Anthropic launched community ambassador program (claude.com)
- LLMs as nudging research towards luke-warm middle (nature.com)
- How do you evaluate a foundation model before you know what it's for? (galsapir.github.io)