The paper introduces Vibe Patenting, an end-to-end testbed where a separately invoked LLM judge evaluates generated patent drafts and provides structured feedback for revision. Across multiple inventions and drafting-agent configurations, judge-guided revision consistently improved judge-assessed quality, while unguided revision tended to saturate.

The useful detail: iterative judge feedback allowed a low-reasoning agent to approach the performance of a substantially more expensive high-reasoning agent. Stronger models and increased reasoning generally improved quality, and domain-specific agentic workflows added further gains. However, validation against a professional patent attorney found meaningful but strongly metric-dependent agreement and systematic calibration differences.