5 Aug 2026, 21:57 UTC125 views6 reactionsread 7 August 2026 Forwarded from @axisofordinaryPhoto
Prime Agent: A self-improving RLM harness for coding and long-running autonomous tasks.
On ARC-AGI-3, it scores 95.5%, surpassing the human-expert baseline, but the gain is not benchmark-specific.
Prime Agent combines three ideas:
1. Recursive Language Models-native programmatic tool calling.
2. Persistent multi-agent orchestration.
3. A self-improving Continual Harness.
Together, they let the model act on its ow…
🔥4👍1🙏1
5 Aug 2026, 07:39 UTC168 views6 reactionsread 7 August 2026 https://arxiv.org/abs/2607.28607v1
👍4👎2
31 Jul 2026, 13:32 UTC320 viewsread 7 August 2026 https://arxiv.org/abs/2607.08716
31 Jul 2026, 13:09 UTC296 views2 reactionsread 7 August 2026 https://github.com/kahip/kahip
🔥2
30 Jul 2026, 09:47 UTC300 views0 reactionsread 7 August 2026 https://arxiv.org/abs/2607.12747
28 Jul 2026, 13:54 UTC329 views4 reactionsread 7 August 2026 https://arxiv.org/abs/2607.22094
👍4
20 Jul 2026, 16:49 UTC492 views6 reactionsread 7 August 2026 File
https://doi.org/10.1093/schbul/sbad168
Cat Ownership and Schizophrenia-Related Disorders and Psychotic-Like Experiences: A Systematic Review and Meta-Analysis (2023)
🎉5👍1
18 Jul 2026, 10:55 UTC483 views5 reactionsread 7 August 2026 https://arxiv.org/abs/2607.12395
💅4❤1
15 Jul 2026, 12:33 UTC517 views9 reactionsread 7 August 2026 https://academic.oup.com/sleep/article/47/1/zsad253/7280269
❤5😱3👎1
15 Jul 2026, 09:52 UTC407 views2 reactionsread 7 August 2026 Forwarded from @pehade_blog
💫 Представляю нарешті нашу статтю як ми робили Лапу!
Тут описано як ми взялись та вирішили 3 проблеми: неефективну токенізацію, брак якісних анотованих даних та відсутність датасетів для слідуванню інструкцій українською. Ми:
1) трансфернули токенізатор завдяки роботі Богдана і Миколи
2) переклад класифікаторів якості і фільтрація претрейну по якості та від дезінформації
3) генерація даних для слідуванню інструкцій
…
👌2
13 Jul 2026, 11:52 UTC334 views7 reactionsread 7 August 2026 Forwarded from @nerdletters
Помилки в мовах програмування це тема для досліджень. Але багато людей користуються ЛЛМ агентами для виконання своєї роботи. Ми довго вчилися робити помилки зрозумілими для людей. Чи це правильна стратегія? Чи потребують агенти іншого формату діагностики про помилки?
Дослідження пана Крішнамурті намагається почати цей діалог. Він сам не каже що дослідження відповідає на всі питання і що є дуже багато обмежень в цьом…
❤3👎3🤔1
12 Jul 2026, 06:16 UTC366 views1 reactionsread 7 August 2026 https://arxiv.org/abs/2606.26493
❤1
Showing the 12 most recent of 20 posts we hold for @natural_origin. View and reaction counts are the latest single reading for each post, not a live figure, and a recent post is still accumulating both. A view count marked ≈ was rounded by Telegram before we ever saw it — t.me prints views in full below 1,000 and to three significant figures above, so ≈1,200,000 means somewhere between 1,150,000 and 1,249,999. Unmarked counts are exact. Text is reproduced from the public post preview and truncated for length.