AI in Action: DeepSeek's Surprising Performance, Platform Engineering ROI, and the Nuances of AI Adoption
Stay updated with the latest in AI and software development, covering DeepSeek's V4-Flash vs. V4-Pro, the true cost of DIY platform engineering, the gap in AI productivity measurement, and real-world AI agent implementations by tech giants.
Welcome to your daily dose of AI and software development insights! Today's digest uncovers surprising model performance, the economic realities of internal platforms, and the crucial distinction between adopting and effectively using AI in engineering workflows. We also look at how major tech companies are leveraging AI agents.
TL;DR
- DeepSeek's V4-Flash model surprisingly outperformed V4-Pro in certain benchmarks, offering a cheaper, faster alternative.
- Building your own platform engineering solution comes with significant, often underestimated, costs beyond just initial development.
- Despite faster AI coding, overall engineering productivity isn't seeing the expected boosts, highlighting a measurement gap.
- Merely adopting AI tools doesn't guarantee effective usage or productivity gains; strategic implementation is key.
- Coinbase, Shopify, and Ramp developed their own AI coding agents but still rely on Anthropic's foundational models.
V4-Flash vs. V4-Pro: DeepSeek promised better and cheaper. It's true, but not how I expected.

A recent analysis of DeepSeek's latest AI models, V4-Flash and V4-Pro, revealed unexpected performance metrics. While DeepSeek aimed to deliver improvements, the results indicate that the V4-Flash model, positioned as a faster and more cost-effective option, often surpassed its V4-Pro counterpart in specific benchmarks. This challenges conventional expectations where 'Pro' versions typically imply superior performance across the board.
The findings suggest that for many use cases, V4-Flash could be the more practical choice, offering a compelling balance of speed and efficiency at a lower price point. This unexpected outcome highlights the importance of thorough benchmarking and understanding specific model strengths rather than relying solely on naming conventions or general performance claims.
The DeepSeek V4-Flash model delivered an unexpected performance edge over V4-Pro in some areas, proving to be a cheaper and faster alternative for certain applications.
Platform Engineering ROI: What it costs to build your own platform - The New Stack

Building an in-house platform engineering solution, while promising significant returns on investment (ROI), comes with substantial costs that extend far beyond initial development. Organizations often underestimate the ongoing expenses related to maintenance, support, and continuous improvement required to keep such a platform effective and competitive. These hidden costs can significantly impact the perceived ROI and long-term viability of a DIY approach.
The article emphasizes the need for a comprehensive financial assessment that accounts for the full lifecycle of a platform. This includes not only engineering salaries and infrastructure but also the less obvious expenditures associated with training, documentation, and the inevitable evolution of technology. A realistic understanding of these costs is crucial for companies deciding whether to build or buy platform engineering capabilities.
The true cost of building your own platform engineering solution often involves significant, overlooked expenses beyond initial development, impacting overall ROI.
AI coding got faster. Why didn't engineering?

Despite the rapid advancements in AI coding tools, which promise to accelerate development cycles, overall engineering productivity has not seen a proportional increase. This discrepancy highlights a critical measurement gap in how companies assess the impact of AI on their software development processes. While AI might generate code faster, other factors in the engineering pipeline, such as debugging, testing, and integration, may not be keeping pace.
The challenge lies in effectively integrating AI-generated code into existing workflows and ensuring its quality and maintainability. Without a holistic approach to productivity measurement and process optimization, the full benefits of faster AI coding remain untapped. The article suggests that organizations need to re-evaluate their metrics to capture the true value (or lack thereof) that AI brings to the entire engineering team.
The acceleration in AI coding has yet to translate into widespread engineering productivity gains, pointing to a need for better measurement and integration strategies.
AI adoption isn’t the same as AI usage

The distinction between AI adoption and actual AI usage is crucial for organizations looking to leverage artificial intelligence effectively. Simply integrating AI tools into a company's tech stack doesn't automatically mean employees are utilizing them to their full potential or that they are generating tangible benefits. Many companies might boast high adoption rates without seeing a corresponding increase in productivity or innovation.
Effective AI usage requires more than just provisioning access; it demands proper training, clear use cases, and a cultural shift that encourages experimentation and integration into daily workflows. Without these elements, AI tools can become expensive, underutilized assets. The article emphasizes that focusing on maximizing usage and demonstrating real-world value is far more important than just tracking initial adoption figures.
Merely adopting AI tools within an organization does not equate to effective usage or guaranteed productivity benefits; strategic implementation and cultural integration are key.
Coinbase, Shopify and Ramp all built their own coding ...

Leading tech companies like Coinbase, Shopify, and Ramp have invested in developing their own sophisticated AI coding agents to enhance developer productivity. This trend underscores a strategic move by enterprises to tailor AI solutions precisely to their internal development needs and workflows. By building proprietary agents, these companies aim to gain a competitive edge and optimize their software development lifecycles.
Interestingly, despite their significant investment in custom AI agents, all three companies continue to pay Anthropic for its foundational models. This highlights a prevalent hybrid strategy in the enterprise AI space: leveraging powerful, general-purpose models from leading AI providers while building specialized applications on top. This approach allows them to benefit from advanced AI research and scale while maintaining control over their specific operational logic.
Major tech players like Coinbase, Shopify, and Ramp are developing proprietary AI coding agents but still rely on Anthropic's foundational models, indicating a hybrid approach to enterprise AI.