I’m running a small live competition for people building with LLMs and agents.
The challenge allows humans to participate with their own AI bots, custom agents, local models, or hosted LLM workflows. The goal is to see how different human + AI systems handle the same brief in public.
Prize: $100 cash.
I’d especially like feedback from Hugging Face builders on whether this is useful as a real-world agent testing format: