← All speakers

Bio, Work & Ideas

Siddharth Ahuja

Siddharth Ahuja is a designer, engineer, and creator of Blender MCP, an open-source integration that lets AI assistants operate professional 3D software through natural language. His work makes sophisticated creative tools accessible to people who know what they want to create but have not mastered complex specialist interfaces.

Ahuja developed Yoguide, a Copenhagen Institute of Interaction Design project that used smartphone-based computer vision to help blind and visually impaired people practice yoga independently. It received a UX Design Award in 2020. He subsequently worked as an entrepreneur-in-residence at LEGO Ventures and, from 2021 to 2024, co-founded and ran Prophecy, a product consultancy whose clients included Loop Health and Headout.

In 2025, Ahuja introduced Blender MCP, applying his experience in design, prototyping, and engineering to Blender’s steep learning curve. Because AI assistants can write code and Blender can execute scripts, he built an integration allowing users to describe scenes, objects, lighting, and animation conversationally.

  • Natural-language 3D creation: Blender MCP connects clients such as Claude and Cursor to Blender through the Model Context Protocol and a custom add-on. Integrations with Sketchfab, Poly Haven, and Hyper3D Rodin enable assistants to retrieve or generate assets and assemble complete scenes.
  • Lean, differentiated agent tools: Ahuja found that overlapping capabilities make models unreliable at selecting the correct action. His approach favors clearly distinct tools organized around the creator’s intended outcome, treating orchestration as a product-design challenge.
  • Cross-application creative orchestration: His Ableton MCP extends the same architecture to music production. Combining both integrations, he demonstrated an assistant creating a dragon scene, adjusting its lighting, and generating a corresponding soundtrack across separate applications.

Ahuja envisions creators directing games, films, and music through their artistic intent while AI coordinates specialized software and external services. He also acknowledges that language models still struggle with three-dimensional spatial reasoning and that generated results remain uneven.

Read the topics behind these talks

1 conference talk

Key ideas

Scroll to read ↓

Blender’s scripting interface lets a language model turn a scene description into operations. Siddharth Ahuja’s experiments show why distinct tools matter, and how the same approach can connect generated objects, animation, and music.

  • Why a donut becomes an interface problem
    0:20 ↗
  • A dragon that gets the brief
    1:57 ↗
  • The model selects tools; Blender executes the work
    3:25 ↗
  • Distinct tools reduce selection ambiguity
    5:47 ↗
  • Generated assets, animation, and image references
    7:44 ↗
  • Terrain nodes, game scenes, and animated materials
    9:04 ↗
  • A Blender racing scene becomes a Runway clip
    10:30 ↗
  • Returning to the donut
    11:27 ↗
  • A workflow organized around creator intent
    12:03 ↗
  • A dragon and its soundtrack
    14:04 ↗
  • What creators still need to direct
    15:15 ↗

References