DeepSeek-V2 shocked the industry with its 236B mixture-of-experts design activating only 21B parameters per token, achieving GPT-4-class performance at 80% less inference cost than the average model. This unlocked API price wars that cut costs by up to 80%, and its open release undercut proprietary rivals like Grok-2 on affordability by 70%.

Comments on "DeepSeek-V2"
Create a free account or sign in to join the discussion.
Sign in to join the conversation