The year-end release
DeepSeek released V3 on December 26, 2024. Its announcement described a mixture-of-experts model with 671 billion total parameters and 37 billion activated parameters, alongside model materials and API access.
The release became another reason for developers to revisit the relationship between model capability, architecture, and serving cost.
Separate three different cost questions
Training a model, hosting its weights, and buying API requests are different activities with different cost structures. A claim about efficiency in one does not directly establish the price or reliability of another.
For a business choosing a service, the useful comparison includes throughput, accepted output quality, operational support, and the effort required to integrate the model. Self-hosting adds hardware planning and ongoing maintenance to that calculation.
Preserve a stable evaluation
Run the same task set against the new candidate and the current system. Include formatting constraints, domain terminology, and incomplete source material, not only general questions. If an API alias changes its underlying model, rerun those checks before assigning the result to a production workflow. Architecture announcements create opportunities, but the application's acceptance criteria should remain the basis of the decision.
Official sources
This article covers an AI industry event. XMH.NET specializes in image generation and editing APIs; coverage does not imply that every model, product, or feature described is available through our service.