DeepSeek V4 Pro vs. Kimi K3: What Months of Real-World AI Agent Work Taught Me This article is based entirely on my personal experience using both models extensively through Pi.dev and my ARAYA framework. It is an independent research exercise, not a sponsored comparison, a reproduction of marketing claims, or an analysis based on technology news. This Is Not a Benchmark Most comparisons between artificial intelligence models begin with public benchmarks, release announcements, pricing tables, or carefully prepared demonstrations. My comparison began somewhere very different: inside real repositories, with incomplete context, conflicting documents, failing gates, distributed agent responsibilities, Git history , runtime evidence, and requirements that could not be considered complete merely because the generated code looked correct. Over several months, weeks, and many hours of intensive use, I worked with DeepSeek V4 Pro and Kimi K3 through Pi.dev while developing and operat...
Data Platform Architecture & AI Engineering
Essays, architecture insights and reflections on data, AI and society