Aleph Alpha releases Kolibri, an open German-English AI model for regulated sectors

Aleph Alpha has released Kolibri, an open-weight language model designed for German and English work in regulated organizations. The company positions the model for public administration, industry, aerospace, banks, and other users that want to operate AI systems on their own infrastructure. Kolibri uses a mixture-of-experts architecture. It contains 78.1 billion parameters in total, but …

Read more

Jev promises to make AI automation faster, cheaper, and more reliable

TypeSafe AI has introduced Jev, a new model designed to make fast, structured decisions for software rather than generate conversational text. The company says Jev can classify, rank, route, score, and select from predefined options in as little as 70 milliseconds. Unlike general-purpose chatbots, Jev does not produce open-ended prose. Developers provide text or structured …

Read more

Ox Alpha steps out of the shadows as Z.ai’s GLM-5.3-Flash

Chinese AI company Z.ai has confirmed that the previously anonymous Ox Alpha model is GLM-5.3-Flash, a new multimodal member of its GLM model family. The company says it plans to release the model’s weights, allowing developers to run, adapt and build products with the system. Luz Ding reports for Bloomberg that Ox Alpha rose rapidly …

Read more

Thomson Reuters launches proprietary AI model

Thomson Reuters has launched Thomson, its first proprietary large language model, and plans to use it first in CoCounsel Legal. Thomson Reuters announces in an official press release that it trained the model in house, using an open source base model and about $40 million in investment for talent and computing resources. The company says …

Read more

Google launches Gemini 3.7 Flash with lower introductory API prices

Google has introduced Gemini 3.7 Flash, a new version of its lower cost AI model aimed at coding, workplace tasks and AI agents that use software tools. The company is offering an introductory API price of $0.75 per million input tokens and $3.75 per million output tokens through the end of 2026. Prices are set …

Read more

Now available: Qwen3.8-Max promises days of autonomous work

Qwen has made its Qwen3.8-Max model generally available through QwenCloud, positioning it as its most capable model so far. Qwen announces in its official blog that the model has 2.4 trillion total parameters, with 95 billion active parameters, and that open weights will follow next week. The company says Qwen3.8-Max is built on the Qwen …

Read more

Thinking Machines releases Inkling-Small with near-flagship performance

Thinking Machines Lab has released Inkling-Small, an open-weights AI model designed to deliver much of the performance of its larger Inkling model with substantially lower computing demands. The company says the model performs similarly to Inkling across reasoning, coding and multimodal tasks while using fewer active parameters for each generated token. Inkling-Small has 276 billion …

Read more

Palmier Pro lets Claude and Codex take on video editing grunt work

Palmier Pro is a new open-source video editor for macOS that combines a traditional timeline with AI generation and controls for AI agents. Co-founder Harrison Tin presents the project in a Show HN post on Hacker News, describing it as a tool designed to reduce repetitive work in video production rather than replace editorial judgment. …

Read more

Google launches Gemini 3.6 Flash and a restricted cybersecurity model

Google is expanding its Gemini model family with two lower-cost general-purpose models and a cybersecurity-focused system that is initially limited to governments and trusted partners. The releases underline Google’s effort to compete on the price, speed and operating efficiency of AI tools, particularly for companies building automated workflows. The most capable of the new public …

Read more

PrismML brings a 27B-parameter AI model to smartphones

PrismML introduces Bonsai 27B, a compressed version of Qwen3.6 27B that the company describes as the first model of its size to run entirely on a smartphone. PrismML announces in an official blog post that the new model shrinks a 27 billion parameter system down to a few gigabytes without abandoning its reasoning, tool use, …

Read more

×