AI News · August 14, 2026

Langfuse improves trace transparency, Gemini 3.6 Flash price halves

270,384 stars today86 new projects+7,887 skills

Langfuse v4.11.0 shipped with a fix that makes observation data honest when it exceeds the load cap, tightening transparency for teams monitoring LLM traces. The release is one of eight minor updates across the ecosystem today, including n8n, Ollama, MinerU, Cline, CrewAI, Pydantic AI, and llama.cpp.

The open-source AI ecosystem gained 270,384 GitHub stars today across the tracked project index, with 86 new projects added. The weekly total for new projects stands at 1,921, alongside 100 new Hugging Face spaces. The top gainer today is "skills" with 7,887 stars.

Releases of the day

Langfuse v4.11.0 addresses trace honesty past load caps, a change that affects users monitoring high-volume pipelines. n8n 2.34.6 fixes the details field in continueErrorOutput mode and migrates Google Ads from API v21 to v25. Ollama v0.32.11 adds support for DeepSeek Harness in its launch command and integrates Meta's Muse Code agentic coding CLI. MinerU 3.4.5 fixes docx tables that were silently dropped when cells contained non-text special characters. Cline v3.0.55 changes auto-update behavior so installations wait while a CLI is attached to the Hub. CrewAI 1.15.16 introduces execution context management with UUID support and records flow-ending exceptions. Pydantic AI v1.107.5 closes a security advisory for DNS rebinding in its local dev web chat UI. llama.cpp b10427 optimizes q4_K dense FFN on Intel GPUs by fusing gate, up, and GLU multiplications.

Langfuse v4.11.0 featureWhat's Changed Features · * feat(trace): tell the truth about observations past the load cap (LFE-14993) by @nkabardin in https://github.com/langfuse/langfuse/pull/16033 *... n8n [email protected] 2.34.6 (2026-08-14) · Bug Fixes · * core: Allow details field in continueErrorOutput mode (#36229) (a5b4822) * Google Ads Node: Migrate from sunset v21 API to v25 (#36260)... Ollama v0.32.11 What's Changed * ollama launch dsh now supports DeepSeek Harness, DeepSeek's open-source agent harness * ollama launch muse now supports Muse Code, Meta's agentic coding CLI *... MinerU mineru-3.4.5-released What's Changed * fix #5357: docx tables silently dropped when cells contain non-text special characters by @hwb96 in https://github.com/opendatalab/MinerU/pull/5358 * fix... Cline cli-v3.0.55 Auto-updates no longer install while a CLI is attached to the Hub. The update is recorded at startup and installed on exit, once the Hub confirms nothing else is attached, so a... CrewAI 1.15.16 What's Changed · Features - Introduce execution context management with UUID support - Record what kind of exception ended a flow - Record when a trace batch is shared with AMP... Pydantic AI v1.107.5 What's Changed 🛡️ Security * GHSA-q2xc-rrxj-58x9: the local dev web chat UI (Agent.to_web(), clai web) didn't validate the Host header, so DNS rebinding from a website you... llama.cpp b10427 sycl: fuse mul_mat(gate) + mul_mat(up) + GLU for q4_K dense FFN (#26779) · Measured on Arc Pro B70 (Battlemage, Level Zero), llama-bench -r 20, two interleaved rounds, tg128: ·...

Pricing moves

Input token prices moved across seven models today. Google cut Gemini 3.6 Flash input price by 50 percent to $0.75 per 1M tokens, with its batch variant down 49 percent to $0.38. Z.AI reduced GLM 5.2 input by 20 percent to $0.39 and GLM 4.6 input by 9 percent to $0.50. Moonshot AI lowered Kimi K2.6 input by 3 percent to $0.56. Alibaba raised Qwen3 Next 80B A3B Instruct input by 11 percent to $0.10 and Qwen3.5-35B-A3B input by 79 percent to $0.25.

▼ 20%GLM 5.2Z.AI · input $0.49 → $0.39 per 1M tokens
▼ 3%Kimi K2.6Moonshot AI · input $0.58 → $0.56 per 1M tokens
▲ 11%Qwen3 Next 80B A3B InstructAlibaba · input $0.09 → $0.1 per 1M tokens
▲ 79%Qwen3.5-35B-A3BAlibaba · input $0.14 → $0.25 per 1M tokens
▼ 9%GLM 4.6Z.AI · input $0.55 → $0.5 per 1M tokens
▼ 50%Gemini 3.6 FlashGoogle · input $1.5 → $0.75 per 1M tokens
▼ 49%Gemini 3.6 Flash (batch)Google · input $0.75 → $0.38 per 1M tokens

Research pick

A Hugging Face community paper by Yuanyang Yin, Gongxuan Wang, Yifan Zhan and others presents Alaya-EVOKE, a method for interactive world models that avoids holding long history in the denoiser context or key-value cache. The approach addresses conflicting demands between persistent memory, responsive interaction, and long-horizon generation.

The bigger picture

Today's releases lean toward infrastructure polish and security fixes, while the price drops from Google and Z.AI continue a downward trend in inference cost. The sharp price increase for Qwen3.5-35B-A3B stands out as an exception.

Source: olud.ai tracking of 10,000+ open-source AI projects, 300+ models and live provider pricing. Figures are measured, not estimated. All releases · Live pricing · Latest in AI

← All AI news