Kaggle is the largest global AI community of developers, researchers, and enthusiasts who compete, collaborate, and benchmark what's next in AI.

San Francisco
Export your AutoScientist open weights to Kaggle 💻 Can't wait to see what the community builds and shares!
Proud to support 33M+ Kaggle builders. 🎉 AutoScientist open weights now export directly to @kaggle.
12
86
10,141
Finish the sentence: "My Kaggle workflow isn't complete without..."
13
45
11,511
Post-train an open model into a coding agent that accelerates developer workflows on everyday hardware. The Gemma 4 Developer Agent Competition, hosted by @googlegemma and Kaggle, is live! 💰 Total Prize Pool: $100,000 ⏰ Entry Deadline: November 25, 2026
14
94
1,243
70,970
How well can frontier models actually code human-preferred motion graphics? 🎬⚡️ We’re excited to introduce @HeyGen's Code2Video Bench on Kaggle, a standardized benchmark measuring how well models can generate deterministic, high-quality @HyperFrames_ code from natural language creative briefs. 💡 The Problem: Static image benchmarks and traditional video generation evaluations don’t measure whether code generation models can follow detailed creative briefs to programmatically control timing, animation sequences, and asset placement over time. 🛠️ How it Works: 1. Creative Brief: Models receive commercial-grade motion design prompts (example: “Create a 6.8s Figma-style product hook that zooms out from a blooming vector flower into a 3-card web page and live timeline editor”). 2. Code to MP4: The model generates HTML/CSS/JS, deterministically rendered into output.mp4 via the open-source Hyperframes engine in isolated Kaggle Harbor sandboxes. 3. The Auto-Judge: HeyGen’s in-house model architecture, trained on human preferences across a 5-axis evaluation framework, evaluates the rendered MP4. Give it the original prompt and two candidate videos, and it returns the probability a human would prefer Video A over Video B. It does this once per axis — five independent verdicts, not one blended score. Here’s what we’ve learned so far 🔽
12
11
103
11,436
Motion is the weakest axis across every model: • Broken timing: Models hit every beat in the brief, but elements fire before previous animations settle. • Mid-flight collisions: Text is laid out for the final frame, crashing along the path it travels. • Off-by-one reveals: Even #2 GPT-6-Astra slips on typewriter reveals, stopping one character short ("Make it your own" holds on screen as "Make it your owr").
2
3
1,595
When unsure how to fill a frame, models play defensive CSS: small type, thin lines, faint backgrounds, wide margins. Ask for a world map filling 2/3 of the canvas, and models render a tiny, quiet version in the center. Everything compiles and stays legible, but nothing has presence.
3
1,489
Can clever engineering on a small budget still beat massive compute on a Kaggle leaderboard, or has scale officially won?
17
6
182
23,519
Mass spectrometry detects thousands of molecules in nature, but most remain unidentified. 🌿 Join the Enveda CASMI 2026 - Molecule ID From Mass Spectra Competition, hosted by @enveda and Kaggle, to build machine learning models that predict 2-D chemical structures from LC-MS/MS spectra. • Prize Pool: $50,000 • Entry Deadline: December 7, 2026
5
6
61
12,671
We’re sitting down with Orbit Wars competitors across the hardware spectrum for our first community podcast episode. What questions do you have for them? 👇
4
4
19
11,124
Introducing ExtractBench on Kaggle Benchmarks with @llama_index. When AI agents rely on schema-guided extraction before human review, one truncated schedule or invented value becomes a wrong payment or decision. ExtractBench evaluates models in workflows based on real-world documents across industries such as supply chain, healthcare, and finance, by measuring whether the system returns: ➣ Missing fields as null instead of inventing a value ➣ Source evidence for each value ➣ Every record of each repeated structure GPT-5.6 Sol currently leads at 91%.
6
12
119
14,094
Ready to build models that support knee MRI interpretation? 🩻🦵 Check out this starter notebook by Pilkwang Kim. It is a great starting point to read the DICOM acquisitions, sample and normalise the MRI slices, and turn a pretrained vision backbone into twelve abnormality predictions. kaggle.com/code/pilkwang/rsn…
2
21
87
13,419
What's one thing you learned on Kaggle? 🌟
26
1
74
17,576
Today, we’re excited to launch Adversarial Customer Service on Kaggle Benchmarks, in partnership with @GertLabs. This benchmark is a two-sided security game: one model plays a bank's support agent holding customer records and a verification policy, the other plays a caller who is secretly either the real customer or an identity thief. The agent has to work out which, from the conversation alone, and then either help or refuse. Explore the leaderboard here: kaggle.com/benchmarks/gert-l…
7
15
131
19,203
I knew I belonged on Kaggle when... (finish the sentence) 👇
14
1
42
11,962