
@_akhaliq
AI research paper tweets, ML @Gradio (acq. by @HuggingFace 🤗) dm for promo ,submit papers here: https://t.co/UzmYN5XOCi
A New Role for Relevance Guiding Corpus Interaction in Agentic Search paper: huggingface.co/papers/2607.24…
HiFi-UMI Learning Deployable Manipulation Policies from High-Fidelity UMI Data Alone paper: huggingface.co/papers/2607.25…
A.X K2 just dropped on Hugging Face Large-Scale Sparse MoE (688B / 33B Active) huggingface.co/skt/A.X-K2
From Proprietary to Open-Source Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search paper: huggingface.co/papers/2607.24…
StateAct Program State, before Pixels, for Long-Horizon Computer-Use Agents paper: huggingface.co/papers/2607.22…
kimi k3 in claude code via hf claude
Kimi K3 is out huggingface.co/moonshotai/Kim…
Apple-π Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence paper: huggingface.co/papers/2607.16…
SLAI T-Rex Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD paper: huggingface.co/papers/2607.20…
Reading and Steering Representations of Materials Science Mechanisms in an Open Weight Language Model paper: huggingface.co/papers/2607.20…
DataFlow-Harness A Grounded Code-Agent Platform for Constructing Editable LLM Data Pipelines
Solar Open2 250B just dropped on Hugging Face huggingface.co/upstage/Solar-…
ABot-World-0 Infinite Interactive World Rollout on a Single Desktop GPU
Apple presents Environment-free Synthetic Data Generation for API-Calling Agents huggingface.co/papers/2607.16…
Motif-3-Beta just dropped on Hugging Face ~314B total parameters / ~13B active per token (sparse MoE) 256K context length (262,144 tokens), natively long-context Sparse routing: 384 experts with 8 activated per token, plus 1 shared expert Multilingual, general-purpose t.co/EqRUYr4OGD
RESOURCE2SKILL Distilling Executable Agent Skills from Human-Created Multimodal Resources
VideoChat3 Fully Open Video MLLM for Efficient and Generalist Video Understanding
thinkingmachines Inkling is now available in claude code via hf claude
Harness Handbook Making Evolving Agent Harnesses Readable,Navigable, and Editable
Read It Back Pretrained MLLMs Are Zero-Shot Reward Models for Text-to-Image Generation
Weak-to-Strong Generalization via Direct On-Policy Distillation
Bonsai 27B now available in claude code via hf claude
prism-ml/Ternary-Bonsai-27B now available in claude code via hf claude with @togethercompute
Scalable Visual Pretraining for Language Intelligence
Long-Horizon-Terminal-Bench Testing the Limits of Agents on Long-Horizon Terminal Tasks with Dense Reward-Based Grading
Video Generation Models are General-Purpose Vision Learners
OPSD-V On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators
Vidu S1 A Real-Time Interactive Video Generation Model
LingBot-World 2.0 (Infinity) is out on Hugging Face interactive world model with: Hour-long generation with zero quality drift Rich actions & events: attack, cast spells, shoot, summon storms Agentic world: a Director Agent drives real-time world evolution 720p/60fps. Playable like a game
LingBot-Video is out on Hugging Face MoE-based video foundation model built for embodied intelligence 30B params, only 3B active at inference Augmented with 70K hours of embodied data on top of large-scale internet video pretraining
LingBot-Video is out on Hugging Face MoE-based video foundation model built for embodied intelligence. 30B params, only 3B active at inference Augmented with 70K hours of embodied data on top of large-scale internet video pretraining
RynnWorld-4D 4D Embodied World Models for Robotic Manipulation
Gemma 4 Technical Report
Wan-Streamer v0.2 Higher Resolution, Same Latency
PixWorld Unifying 3D Scene Generation and Reconstruction in Pixel Space
Vision Pretraining for Dense Spatial Perception
bottlecapai/ThinkingCap-Qwen3.6-27B Capability of Qwen3.6-27B with 50% less thinking tokens on average, and over 90% less in best cases. Achieved via finetuning Qwen3.6-27B (Qwen Team, 2026) with state-of-the-art finetuning algorithms on a curated set of problems of various domains and difficulty
The Mirage of Optimizing Training Policies Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning
Program-as-Weights A Programming Paradigm for Fuzzy Functions
I use glm 5.2 in claude code via hf claude almost daily now moved over completely to open models x.com/zRdianjiao/sta…
PerceptionRubrics Calibrating Multimodal Evaluation to Human Perception
CausalMix Data Mixture as Causal Inference for Language Model Training
LiteResearcher A Scalable Agentic RL Training Framework for Deep Research Agent
Orca The World is in Your Mind
open-fusion in claude code with hf-claude
OSWorld2.0 Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks
Ornith-1.0-35B is now available in claude code through hf-claude
LongCat-2.0 dropping on Hugging Face soon
PhysisForcing Physics Reinforced World Simulator for Robotic Manipulation
DiffusionBench On Holistic Evaluation of Diffusion Transformers
baidu/Unlimited-OCR is now number 1 model on huggingface
VISReg Variance-Invariance-Sketching Regularization for JEPA training