Chronicling the Seattle and Pacific Northwest startup scene. SEE MORE by Jon Turow on Jan 27, 2025 at 11:05 am January 27, 2025 at 1:53 pm Editor’s note: This post first appeared on Jon Turow’s ...
DeepSeek looking to regain momentum after being surpassed by Chinese rivals V4-Flash about 105 times cheaper than Anthropic's Claude Fable 5, tests show On performance, ranks on par with Google's ...
DeepSeek-V4.1-Flash is available now on Baseten Model APIs, Baseten announced on September 11, 2026, bringing the 552B-parameter multimodal mixture-of-experts (MoE) model, which pairs 8B active ...
DeepSeek's V4.1 Flash model packs 552 billion parameters at a fraction of rivals' costs, sending competitor stocks tumbling over 8% on launch ...
It’s been almost a year since DeepSeek made a major AI splash. In January, the Chinese company reported that one of its large language models rivaled an OpenAI counterpart on math and coding ...
Brian Wang is a Futurist Thought Leader and a popular Science blogger with 1 million readers per month. His blog Nextbigfuture.com is ranked #1 Science News Blog. It covers many disruptive technology ...
Delivering higher efficiency and reduced KV cache consumption, this new open-source model outperforms several flagship offerings on major coding and reasoning benchmarks.
Researchers at Together AI and Agentica have released DeepCoder-14B, a new coding model that delivers impressive performance comparable to leading proprietary models like OpenAI's o3-mini. Built on ...
This aerial photograph taken on January 14, 2026 shows the building housing the headquarters of Chinese AI startup DeepSeek in Hangzhou, in China's eastern Zhejiang province. Jade GAO/AFP via Getty ...
Even as the geopolitical conversation around AI continues to grow more fraught following the U.S. government's actions to limit the new models from Anthropic and OpenAI, Chinese open source darling ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results