
@intology
Automating the process of discovery.
@intology recognized at this year's Smart AI 100 Summit!
Ron Arel@ronusedh·Thank you to @gracegongGG for honoring @intology at the Smart AI 100 Awards. Grateful for the panels with leaders from @Cisco , @ServiceNow, @nebiusai, @cerebras, @Atlassian, @datadoghq, @Zoom, @asana, and more. Really enjoyed the conversations with: Jeetu Patel · Marc Boroditsky · Julie Choi · Amit Zavery · Abbas Haider Ali · Emilio E. · Brendan Ittelson · Arnab Bose Such a great conference, excited to have been a part of it.
Locus is now verified on the official PostTrainBench leaderboard! Thank you to @hrdkbhatnagar, @full__rank, and @maksym_andr for reviewing and uploading our solutions on the 10-hour setting.
Full leaderboard: posttrainbench.com
Locus is now verified on the official PostTrainBench leaderboard! Thank you to @hrdkbhatnagar, @full__rank, and @maksym_andr for reviewing and uploading our solutions on the 10-hour setting.
Leaderboard: posttrainbench.com
Thank you @jackclarkSF for covering our work with Locus! "AI startup Intology, whose goal “is to automate R&D”, has released a new version of Locus, software it has developed to turn LLMs into capable researchers. The new version of Locus is able to get a score of 44.7% on PostTrainBench, a benchmark which sees how well AI systems can take an open weight model and improve its performance above its baseline." "Posts like this highlight how we are under-eliciting today’s AI systems for their ability to automate AI R&D - especially striking is how the company can jump the performance of Opus 5 by 10 absolute percentage points"
Ron Arel@ronusedh·Thanks to @jackclarkSF for covering @intology's work on Locus! "Posts like this highlight how we are under-eliciting today’s AI systems for their ability to automate AI R&D - especially striking is how the company can jump the performance of Opus 5 by 10 absolute percentage points" Full blog: t.co/JeKSHTB1xP
Our blog: intology.ai/blog/scaling-a… Jack's: importai.substack.com/p/import-ai-46…
On building / no building copilots w/ @tbpn
Ron Arel@ronusedh·More clips from our @tbpn segment. On building / not building copilots.
Thank you to @tbpn for having us on last week. Ron Arel (@ronusedh, co-founder of @intology) spoke on our research direction — our past work, and our new state-of-the-art results on automating post-training.
Intology@intology·The models are improving the models. Locus, our automated AI research system, is SOTA on PostTrainBench and post-trains Qwen3 base models that surpass the human post-trained Qwen3 model. Today, LLMs post-trained end-to-end by Locus are in production to millions. 🧵👇 PostTrainBench evaluates agents' ability to post-train models on various domains given 10 H100 hours. We extend PostTrainBench via PostTrainBench+, which has a greatly expanded compute budget that provides clearer signal on automated post-training capabilities. We find that thousands of H100 hours help distinguish methods' performance post-training Qwen3 1.7B-Base models, and that Locus scales best. In this setting, modes trained by Locus collectively surpass the perforamce of the offical human post-trained Qwen3 1.7B model. In a test of generalization, we ran Locus on all live Kaggle competitions with prize money and public leaderboards. After 16 days, Locus achieved the 4th highest average rank among all participants.
Ron Arel@ronusedh·More clips from our @tbpn segment. On building / not building copilots.
We had a great time presenting at the AI Scientist Summer Workshop in Boston. Thank you to @WengongJin, @YuanqiD, @YEktefaie, @xiwei393, @BotaoYu24, Yikun Zhang for organizing this amazing event, and thank you to @MSFTResearch for hosting us! x.com/intology/statu…
@WengongJin @YuanqiD @YEktefaie @xiwei393 @BotaoYu24 @MSFTResearch And thank you to @AdaFang_ for inviting us and making this all happen.
We will be presenting at the AI Scientist Summer Workshop in Cambridge, this coming Tuesday, August 4, 2026. Ron Arel (@ronusedh) will be presenting. Thank you @AdaFang_ for coordinating! x.com/intology/statu…
Ron Arel (@ronusedh), co-founder of Intology will be on @TBPN today to discuss Locus' automated post-training capabilities, and more x.com/tbpn/status/20…
The models are improving the models. Locus, our automated AI research system, is SOTA on PostTrainBench and post-trains Qwen3 base models that surpass the human post-trained Qwen3 model. Today, LLMs post-trained end-to-end by Locus are in production to millions. 🧵👇 PostTrainBench evaluates agents' ability to post-train models on various domains given 10 H100 hours. We extend PostTrainBench via PostTrainBench+, which has a greatly expanded compute budget that provides clearer signal on automated post-training capabilities. We find that thousands of H100 hours help distinguish methods' performance post-training Qwen3 1.7B-Base models, and that Locus scales best. In this setting, modes trained by Locus collectively surpass the perforamce of the offical human post-trained Qwen3 1.7B model. In a test of generalization, we ran Locus on all live Kaggle competitions with prize money and public leaderboards. After 16 days, Locus achieved the 4th highest average rank among all participants.
Locus is state-of-the-art on PostTrainBench. PostTrainBench gives agents 10 H100-hours to post-train open-weight models on seven benchmarks spanning domains from healthcare to coding. The benchmark uses the performance of the existing human-post-trained models on each of the individual benchmarks as a reference.
Our Artificial Scientist, Locus, achieved a World Record on the NanoGPT speedrun via a fused triton kernel! x.com/intology/statu…
When coding agents consider algorithmic work, they rarely succeed. Instead, they either reason themselves away or regress performance. For example, Autoresearch repeatedly considered reducing the number of value embeddings from 3 to 2, but avoided the change after deeming it risky without any experimentation.
With a high human bar established over a long period of time, built-in contamination prevention, and optimized initialization to reduce the effect of low-hanging fruit, NanoGPT-Bench elicits a clear signal for AI R&D.