Topic: qwen

14 stories found

Friday, September 4, 2026

releases48

ggml/llama.cpp releases: v0.4.0

Version 0.4.0 of llama.cpp was released, adding support for Qwen3.8-Flash-Next and Nemotron-3-Puzzle models, along with several new features like on-demand tensor reading and video input options, making it more versatile for AI language tasks. This update is significant as it enhances the model's capabilities and flexibility, catering to a broader range of applications in natural language processing.

github.comโ†—
open_source30

Show HN: Sageling - a local AI agent for Mac, Qwen 3.5 9B in-process via MLX

Sageling is a new local AI assistant designed for macOS, emphasizing privacy by processing data locally rather than sending it to remote servers. This development matters because it addresses growing concerns over data privacy and control, offering users a more localized and secure alternative to cloud-based AI services.

sageling.aiโ†—

Thursday, September 3, 2026

trending59

Qwen 3.8 27B available on Cerebras at 1500 tokens/s

Qwen 3.8 27B, a large language model, is now accessible on the Cerebras platform with processing capabilities at 1500 tokens per second, enhancing its utility for real-time applications and research. This development matters because it expands the model's reach and performance, potentially accelerating innovation in natural language processing.

inference-docs.cerebras.aiโ†—

Wednesday, August 26, 2026

trending59

Qwen3.8-Flash-Next

Qwen3.8, an updated version of the Qwen AI model, was released with enhanced features to improve natural language processing capabilities. This update is significant as it aims to provide more accurate and contextually relevant responses, potentially advancing the state of AI in text generation and understanding.

qwen.aiโ†—

14 of 14 items shown. Sources: 107 days indexed.