The AI community building the future. hf.co/careers

NYC and Paris and 🌏
Hugging Face retweeted
ok what the fuck. this was a dry google doc a moment ago gave Opus 5.5 my Open Alignment brainstorm gdoc and asked for a video. now I want one for every doc I’ve ever written hahaha
32
14
283
23,905
Hugging Face retweeted
Super happy to release SmolDataEnvs: 5,000 verifiable RL environment tasks for hill-climbing small models in code and data science by @adithya_s_k 100% open source: environments, evals, training! huggingface.co/datasets/Fine…
46
47
446
32,059
Hugging Face retweeted
Releasing SmolDataEnvs πŸ€— 5K+ Verifiable RL Environment tasks for hill-climbing small models in code and data science. Completely open source: environments, evals, training
24
51
478
48,046
Hugging Face retweeted
Today, we open-source Pruna-Qwen-Image-2.1, a set of a few-step LoRA adapters that make Qwen-Image-2.1 by @Alibaba_Qwen up to 6.3Γ— faster for image generation and editing. - π—›π—Όπ˜„ π—²π—³π—³π—Άπ—°π—Άπ—²π—»π˜ π—Άπ˜€ π—Άπ˜? Generate or edit images in just 5 or 8 steps instead of 40. Choose 5 steps for maximum speed or 8 steps for the best balance. - π—›π—Όπ˜„ 𝗴𝗼𝗼𝗱 π—Άπ˜€ π—Άπ˜? The 8-step adapter is our recommended default. It supports text-to-image and single- or multi-image editing at 1K resolution, with up to three reference images. - π—›π—Όπ˜„ π—±π—Όπ—²π˜€ π—Άπ˜ π˜„π—Όπ—Ώπ—Έ?The LoRA adapters load directly on top of Qwen-Image-2.1, while the pipeline remains unchanged. Each adapter uses its own optimized sigma schedule and runs without CFG. π˜›π˜©π˜ͺ𝘴 π˜ͺ𝘴 𝘫𝘢𝘴𝘡 𝘰𝘢𝘳 𝘧π˜ͺ𝘳𝘴𝘡 𝘳𝘦𝘭𝘦𝘒𝘴𝘦, 𝘸π˜ͺ𝘡𝘩 𝘦𝘷𝘦𝘯 𝘣𝘦𝘡𝘡𝘦𝘳 𝘷𝘦𝘳𝘴π˜ͺ𝘰𝘯𝘴 𝘰𝘯 𝘡𝘩𝘦 𝘸𝘒𝘺. πŸš€ Want to try the fast open-source integration? Check it here on Diffusers: buff.ly/1CpiRPp Want to try the fastest image generation and editing endpoint? Check P-Image-Ideogram and P-Image-Edit here on API: buff.ly/gw5tvUV
31
94
906
79,443
Hugging Face retweeted
Alert: Apple just dropped a new model on Hugging Face. It's a Qwen3.5-9B finetune that turns long documents into small page images to save tokens, then pulls up the full text of only the pages relevant to your question πŸ’‘ huggingface.co/apple/LensVLM…
74
276
3,452
274,927
Hugging Face retweeted
day 0 transformers πŸ€— support, have fun building with it!
When several people talk at once, a transcript can get messy fast. Our new Nemotron 3 Diarization model tracks who spoke when, even when voices overlap. It handles up to eight speakers, has 100M parameters, and is now available on @huggingface πŸ€—
2
7
46
16,468
Hugging Face retweeted
When several people talk at once, a transcript can get messy fast. Our new Nemotron 3 Diarization model tracks who spoke when, even when voices overlap. It handles up to eight speakers, has 100M parameters, and is now available on @huggingface πŸ€—
189
854
9,628
1,001,395
Hugging Face retweeted
Introducing FLUX 3 Action. An open weights 7B World Action Model that achieves first place on the RoboLab benchmark. It outperforms the previous best open model by 6.1 percentage points while using 56% fewer parameters and running up to 3.95x faster.⁠⁠ FLUX 3 Action removes the usual trade-off between world action model performance and VLA speed: it still predicts video and actions together, but plans more than twice as far ahead and runs faster per second of robot motion than the strongest open VLA. Teams can fine-tune FLUX 3 Action on their own demonstrations to create policies for a particular robot and task. Together with @nvidia, we also integrated FLUX 3 Action natively into @huggingface's LeRobot, with fine-tuning recipes included and edge deployment on NVIDIA Jetson. Beyond robotics, we’re also seeing promising results training task-specific policies for acting in simulated environments like gaming, controlling a vehicle, computer use, and wherever else a model needs to understand a visual environment and then choose what to do next. FLUX 3 Action builds on the same image, video, and audio pretraining as FLUX 3, but uses a smaller architecture designed for practical deployment. In midtraining, we trained the model to predict actions and future frames together. We’re releasing the weights, code, fine-tuning recipe, benchmarks, and reproducible examples so researchers and developers can build on the model with their own robots, environments, and tasks (see below).
75
234
1,785
164,019
Hugging Face retweeted
We’re proud to sponsor Open Together, a free community event hosted by @huggingface on Friday, October 16, to kick off Open Source AI Week in SF. The evening is split into two parts: ➑️ 6PM – 9PM: 36 live community demos, food, drinks, and time to connect with open-source builders πŸͺ© 9PM – 12AM: Full dance floor with live DJ sets Doors open at 6 PM, and the first 500 people through the door get collectible HF swag. πŸ€— RSVP: luma.com/opentogether
4
13
54
13,822
Hugging Face retweeted
Voice agents still don’t understand who’s speaking to them. That’s a huge gap compared with humans, hidden by all the β€œphone-call” demos. But that changes today! NVIDIA is open-sourcing Nemotron 3 Diarization: a model that can reliably track speakers in live conversations, under a commercial-friendly license! In my tests, the quality is really good with one-second speech chunks. So we can use it for voice agents! I tested it with Reachy Mini and speech-to-speech running on a DGX Spark. It’s super fun to see the robot notice a new voice, ask for a name, and remember it. The model has day-zero integration with Transformers! Kudos to the NVIDIA team for shipping useful tools for the whole community!
41
53
521
75,053
Hugging Face retweeted
I'm happy to announce that I've joined Hugging Face. What started as a personal project back in February is now something I get to work on full time. Local AI has grown explosively this year. I've said this since the early oMLX releases: I want my friend who bought a MacBook yesterday to be able to run AI on it today. I believe MLX has that potential. Apple Silicon is the easiest entry point for a regular person to get started with AI, with no complicated hardware to assemble. And the open source community, including Hugging Face, has been growing that potential. That support has already made a huge difference. Hugging Face is the best place for me to support oMLX and the MLX community with everything I have. For anyone getting into local AI, the first step usually starts at Hugging Face. Mine did too. I'm proud that I get to work at that entry point, where new ideas can be tried out. oMLX stays exactly where it is, under the same Apache 2.0 license in the same repository, and I'll keep leading the project, same as before. What changes is that I can now spend far more time on it, move faster, and build something sustainable for the long term together with all the contributors who have put so much into it. I also want to do more for MLX as a whole. oMLX is built on top of transformers, mlx-lm, mlx-vlm and the rest of the ecosystem, and I'm grateful to the people behind them. Rather than keeping everything inside oMLX, over time I want to push work upstream where it makes sense. And where the community needs something that doesn't exist yet, oMLX is a good place to try it first. Thank you to the 264 contributors who have built oMLX with me, and to everyone who filed issues with detailed logs and reproductions so we could fix things. oMLX would not be what it is without you. oMLX continues in the same place, in the same way. Just much faster. HF's announcement: huggingface.co/blog/omlx
130
101
1,103
92,665
Hugging Face retweeted
Run GGUF models directly with transformers. This work brings ggml's Metal kernels to the transformers ecosystem, increasing compatibility and performance. More info below
Millions of GGUF downloads later, those same llama.cpp checkpoints can now run in πŸ€— transformers. Same models, more ways to use them, and fast local inference on Mac powered by ggml kernels! Blog: huggingface.co/blog/transfor… ggml kernels: huggingface.co/ggml-org/kern…
21
51
394
51,829
Hugging Face retweeted
How do you keep @vllm_project moving at the speed of light without excluding users who run diverse models on diverse hardware? In a new PyTorch Foundation blog, contributors from @IBM, @Meta, and @huggingface introduce hardware-agnostic layers designed to balance frontier performance with portability, helping ensure vLLM continues to meet the needs of the broader open source ecosystem. Read the blog to learn more: bit.ly/4xEH7Wk @hmellor_ @th_ortner
4
9
57
14,481
Hugging Face retweeted
Introducing the Decision Index 0.1 βš–οΈ a rigorous leaderboard comparing jev with 30+ open weights decision models 35+ benchmarks. asking 130K questions to each model testing knowledge 🧠, automation βš™οΈ, understanding πŸ€”and even creativity 🎨 huggingface.co/spaces/multim…
28
52
312
42,586
Hugging Face retweeted
A great blog written by @loldedxd & @ariG23498 πŸ€—πŸ‘ huggingface.co/blog/ariG2349… funes, by @huggingface, turns past agent sessions into memory your agents can actually use. It indexes Claude Code, Codex, pi, and Hermes traces into one local Lance dataset, then gives the agent 'recall' and 'get' tools. The next time a task depends on old reasoning, the agent can pull the original passage back. No LLM summarizing your traces at ingest.
11
5
53
14,495
Hugging Face retweeted
Training models is becoming easier and easier - just look at this and TRL - especially with agents! You're missing out if you're still using off the shelf models for all your tasks!
Introducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of stock TRL with less peak memory, while models stay in their native HuggingFace format. Star us on GitHub: github.com/whitecircle/halo
35
113
1,342
175,753
Hugging Face retweeted
MiMo-V2.6-Pro debuts as the top open weights model on the Artificial Analysis Intelligence Index (46). At $0.13 per Intelligence Index task, it lands on the Intelligence vs. Cost per Task Pareto frontier @Xiaomi has just released MiMo-V2.6-Pro, an open weights model with major advances in intelligence over its predecessor, MiMo-V2.5-Pro (Intelligence Index: 26). Despite the improvement, it retains the same attractive pricing at $0.435 per 1M input tokens (with a 99% cache-hit discount) and $0.87 per 1M output tokens. This makes MiMo-V2.6-Pro one of the most cost-efficient models to deploy. MiMo-V2.6-Pro is an MoE model with 1.02T total parameters and 42B active parameters. Stay tuned for additional analysis of the model. Check out MiMo-V2.6-Pro full benchmarking breakdown here: artificialanalysis.ai
13
20
308
40,843
Hugging Face retweeted
The number one trending model on HF is an open-source multilingual system 1 decision model, just a few days after Jev started trending. The open-source AI community is awesome!
78
177
2,624
136,883
Hugging Face retweeted
leaving a few beginner friendly guides for classifiers (normal and zero-shot) as well as where you can find them on the Hub opt for DeBERTa and ModernBERT ones > huggingface.co/models?pipeli… > huggingface.co/tasks/zero-sh… > multimodal (image <> text) huggingface.co/docs/transfor…
people who compare Jev against GPT-5.6 has never fine-tuned BERTForXYZ for living and it shows joke aside I always found zero shot classifiers to be fascinating and was sad we never got people to scale them as much as decoder only models many problems solved with LLMs could have been solved with them, it was a skill issue
9
44
387
43,091