Google launches Gemini 3.7 Flash with lower introductory API prices

Google has introduced Gemini 3.7 Flash, a new version of its lower cost AI model aimed at coding, workplace tasks and AI agents that use software tools. The company is offering an introductory API price of $0.75 per million input tokens and $3.75 per million output tokens through the end of 2026. Prices are set …

Read more

Now available: Qwen3.8-Max promises days of autonomous work

Qwen has made its Qwen3.8-Max model generally available through QwenCloud, positioning it as its most capable model so far. Qwen announces in its official blog that the model has 2.4 trillion total parameters, with 95 billion active parameters, and that open weights will follow next week. The company says Qwen3.8-Max is built on the Qwen …

Read more

Thinking Machines releases Inkling-Small with near-flagship performance

Thinking Machines Lab has released Inkling-Small, an open-weights AI model designed to deliver much of the performance of its larger Inkling model with substantially lower computing demands. The company says the model performs similarly to Inkling across reasoning, coding and multimodal tasks while using fewer active parameters for each generated token. Inkling-Small has 276 billion …

Read more

Palmier Pro lets Claude and Codex take on video editing grunt work

Palmier Pro is a new open-source video editor for macOS that combines a traditional timeline with AI generation and controls for AI agents. Co-founder Harrison Tin presents the project in a Show HN post on Hacker News, describing it as a tool designed to reduce repetitive work in video production rather than replace editorial judgment. …

Read more

Google launches Gemini 3.6 Flash and a restricted cybersecurity model

Google is expanding its Gemini model family with two lower-cost general-purpose models and a cybersecurity-focused system that is initially limited to governments and trusted partners. The releases underline Google’s effort to compete on the price, speed and operating efficiency of AI tools, particularly for companies building automated workflows. The most capable of the new public …

Read more

PrismML brings a 27B-parameter AI model to smartphones

PrismML introduces Bonsai 27B, a compressed version of Qwen3.6 27B that the company describes as the first model of its size to run entirely on a smartphone. PrismML announces in an official blog post that the new model shrinks a 27 billion parameter system down to a few gigabytes without abandoning its reasoning, tool use, …

Read more

China’s Kimi K3 rivals the world’s biggest AI models and it’s open source

Moonshot AI, the Beijing based startup backed by Alibaba, has released Kimi K3, a 2.8 trillion parameter model the company calls in an official blog post the world’s first open 3T-class model. The system combines two new architectural components, Kimi Delta Attention and Attention Residuals, with a 1 million token context window and native vision …

Read more

Thinking Machines shows Inkling, an open source multimodal AI model

Thinking Machines, the AI startup founded by former OpenAI chief technology officer Mira Murati, has released its first major language model under an open source license. Carl Franzen reports for VentureBeat that the model, called Inkling, targets enterprises that want to run AI on their own servers while keeping costs low and avoiding ideological filtering. …

Read more

“Project Moonraker”: Amazon invests heavily in Alexa’s brain upgrade

Amazon is developing a new Alexa project, codenamed Moonraker, designed to let its voice assistant complete several tasks from a single request. Eugene Kim reports for Business Insider that internal planning documents describe the project as enabling “multi-request” engagements, such as booking a ride and texting a friend in one interaction. Moonraker builds on Alexa+, …

Read more

×