⚡Trending:
L
LiveKit@livekit·12h ago
🛠️ Tooling

LiveKit Agents 1.8.3: Gemini 3.8 Live NON_BLOCKING default, GPTLive Azure, AssemblyAI Universal, safer tool handoffs

LiveKit published Python [email protected] on GitHub (2026-09-26; PyPI livekit-agents 1.8.3): Gemini 3.8 Live defaults `tool_behavior=NON_BLOCKING` and sends FunctionTool schemas as `parameters_json_schema` on the text API; OpenAI gains `GPTLiveModel.with_azure` and moves the Realtime inference client into agents core; STT adds AssemblyAI `universal-3-6-pro` plus `language_confidence`, and configurable ElevenLabs realtime chunk duration; voice fixes prevent parallel tools from canceling handoffs, pair tool outputs by call_id, forward frames as-is in fallback/stream adapters, and fix LiveAvatar interruption/plugin registration. Same-day JS `@livekit/[email protected]` deepens agent_turn telemetry and adaptive interruption.

LiveKit Agents 1.8.3: Gemini 3.8 Live NON_BLOCKING default, GPTLive Azure, AssemblyAI Universal, safer tool handoffs
⚡ Key Takeaways
  • •Shipped: [email protected] / PyPI 1.8.3; JS @livekit/[email protected] same day
  • •Gemini 3.8 Live: default NON_BLOCKING tool_behavior; FunctionTool via parameters_json_schema on text API
  • •OpenAI: GPTLiveModel.with_azure; Realtime inference client moved into agents core
  • •STT: AssemblyAI universal-3-6-pro + language_confidence; configurable ElevenLabs realtime chunk duration
  • •Voice correctness: parallel tools no longer cancel handoffs; pair tool outputs by call_id; adapters forward frames as-is
Read details→
C
Cline@cline·19h ago
🛠️ Tooling

Cline Desktop v0.0.37: About & What's new surfaces; Bedrock OpenAI inference profiles fix reasoningConfig errors

Cline Desktop v0.0.37 (2026-09-26) adds Settings → About (version/channel, check/restart updates, recent release notes, report issue), a one-time What's new dialog after updates, macOS Help→Export Diagnostics, and fixes Bedrock OpenAI inference-profile calls that failed with Unknown parameter reasoningConfig when reasoning effort was set.

Cline Desktop v0.0.37: About & What's new surfaces; Bedrock OpenAI inference profiles fix reasoningConfig errors
⚡ Key Takeaways
  • •Shipped: desktop-v0.0.37 vs desktop-v0.0.36 on GitHub Releases
  • •UX: Settings About page with version/channel, update buttons, release notes links, Report an issue
  • •Onboarding: one-time What's new after update; Highlights on About reopens it
  • •Ops: macOS Help→Export Diagnostics opens the export directly
  • •Bedrock: OpenAI models via us.openai./global.openai. inference profiles no longer fail on reasoningConfig when reasoning effort is set
Read details→
C
Cline@cline_dev·19h ago
🛠️ Tooling

Cline Desktop v0.0.37 Released: Adds Built-in Updater & Highlights Dialog, Patches Bedrock Reasoning Config

Cline Desktop (@cline_dev) has shipped release v0.0.37. The update introduces a comprehensive lifecycle management experience, embedding a dedicated 'About' view within Settings that enables in-app update checks, one-click restarts, and inline changelog inspection. Users are greeted with an updated 'What's New' dialog highlighting recent architectural milestones including SSH remotes, Git worktrees, pull request telemetry, and parallel sub-agents. Crucially, the release patches an AWS Bedrock inference bug that triggered 'Unknown parameter: reasoningConfig' on OpenAI models, while adding a macOS diagnostic export shortcut.

Cline Desktop v0.0.37 Released: Adds Built-in Updater & Highlights Dialog, Patches Bedrock Reasoning Config
⚡ Key Takeaways
  • •Introduces an in-app About settings pane featuring direct update checking, automated restart patching, and release notes.
  • •Launches a unified What's New dialog summarizing SSH remotes, Git worktrees, PR composer status, and parallel sub-agents.
  • •Fixes an AWS Bedrock regression causing OpenAI models to crash with 'Unknown parameter: reasoningConfig'.
  • •Adds a direct macOS Help menu shortcut for one-click diagnostic log exporting.
  • •Signed binaries for macOS, Windows, and Linux available directly on GitHub Releases.
Read details→
V
Vercel AI SDK@vercel·19h ago
🛠️ Tooling

@ai-sdk/perplexity 5.0.0: breaking migrate from Sonar Chat Completions to Perplexity Agent API

Vercel AI SDK shipped @ai-sdk/[email protected] (2026-09-26), a breaking migration from Sonar Chat Completions to Perplexity Agent API. Replace Sonar model IDs and provider options with Agent API presets, models, and tools; request/response metadata, raw stream events, and usage/cost shapes change; Sonar PDF input and image/video results are unsupported. Follow-up fixes recover terminal text, accept native tool traces (e.g. finance), and dedupe URL citation sources while preserving search-result IDs.

@ai-sdk/perplexity 5.0.0: breaking migrate from Sonar Chat Completions to Perplexity Agent API
⚡ Key Takeaways
  • •Shipped: @ai-sdk/[email protected] on GitHub Releases / npm 5.0.0
  • •Breaking: language generation moves to Perplexity Agent API; rewrite Sonar model IDs and provider options to presets/models/tools
  • •Dropped: Sonar PDF input and image/video results; usage/cost and raw stream metadata shapes change
  • •Streaming: recover missing text from terminal events/output items (incl. incomplete) without duplicating deltas
  • •Citations: emit each source URL once while keeping search-result IDs; accept native Agent API tool traces (e.g. finance) without web-search validation
Read details→
ADSponsored
v
vercel@vercel·19h ago
🛠️ Tooling

Vercel AI SDK @ai-sdk/openai 4.0.78: GPT-6 Sol/Luna allow reasoningEffortUpdate 'none'

Vercel AI SDK shipped `@ai-sdk/[email protected]` on 2026-09-26: request-level and positioned system-message `reasoningEffortUpdate` now accept `'none'` so GPT-6 Sol/Luna conversations can turn reasoning off mid-thread. Updates are validated against each model’s supported efforts—unsupported request-level updates warn and omit; unsupported historical/message-level updates reject. Companion `@ai-sdk/[email protected]` and `@ai-sdk/[email protected]` pick up the dependency.

Vercel AI SDK @ai-sdk/openai 4.0.78: GPT-6 Sol/Luna allow reasoningEffortUpdate 'none'
⚡ Key Takeaways
  • •Shipped: @ai-sdk/[email protected] / npm 4.0.78
  • •Feature: reasoningEffortUpdate 'none' for GPT-6 Sol/Luna (#21518)
  • •Validation: warn+omit unsupported request-level; reject unsupported historical updates
  • •Companions: Azure 4.0.82, Bedrock 5.0.97; see OpenAI provider docs
  • •Install: npm i @ai-sdk/[email protected]
Read details→
P
Pipecat@pipecat_ai·19h ago
🛠️ Tooling

Pipecat 1.12.0: cross-vendor IPA pronunciation transforms, MOQTransport client reconnect, Classifier/LLMClassifier eval surface

Pipecat v1.12.0 (2026-09-26) adds TTSService.pronunciation_transform_ipa() for reusable word→IPA maps across Cartesia/ElevenLabs/Inworld/Deepgram, client-mode MOQTransport reconnect with session-ending markers, Pipecat Classifiers + LLMClassifier, frame-level interruptible, and response_schema on LLM run_inference. PyPI pipecat-ai 1.12.0 is live.

Pipecat 1.12.0: cross-vendor IPA pronunciation transforms, MOQTransport client reconnect, Classifier/LLMClassifier eval surface
⚡ Key Takeaways
  • •Shipped: GitHub v1.12.0 = PyPI pipecat-ai 1.12.0; docs docs.pipecat.ai
  • •Pronunciation: pronunciation_transform_ipa maps words to IPA once; vendors emit phoneme/SSML/slash IPA—register last in text_transforms
  • •Transport: MOQTransport client reconnect within connection_timeout (default 60s); session-ending distinguishes hangup vs outage; relay_url full URL
  • •Eval: Classifiers (YesNo/Choice/Score) + LLMClassifier batch questions; allow_continue=false for final-answer judges
  • •Inference: run_inference response_schema enforced by OpenAI/Anthropic/Google; frames interruptible by default; UninterruptibleFrame deprecated
Read details→
Q
QwenLM@Alibaba_Qwen·20h ago
🛠️ Tooling

Qwen Code 0.24.6: Managed Runtime v2, native Advisor tool, /batch-api, Hosted & Durable Sessions

Alibaba’s open-source terminal coding agent Qwen Code shipped v0.24.6 on 2026-09-26 (npm @qwen-code/qwen-code 0.24.6; Desktop desktop-v0.24.6 and SDK TypeScript sdk-typescript-v0.1.16 same day): CLI declares Managed Runtime v2 execute/status/cancel and mounts tool ops on the worker; adds a native Advisor consultation tool with usage limits; `/batch-api` for agent-prepared batch workflows; serve gains guarded Hosted Runtime, Spring control plane + dual-path WebShell, and model/reasoning selection at session creation; core adds Durable Managed Session journal and failover. Release notes list no known breaking changes.

Qwen Code 0.24.6: Managed Runtime v2, native Advisor tool, /batch-api, Hosted & Durable Sessions
⚡ Key Takeaways
  • •Shipped: v0.24.6 / npm 0.24.6; Desktop desktop-v0.24.6 same window
  • •Runtime: Managed Runtime v2 execute/status/cancel contract; tool ops mounted on the worker
  • •Advisor: native advisor tool with consultation behavior and usage limits
  • •Batch: /batch-api agent-prepared Batch API workflow
  • •Hosted: guarded Hosted Runtime + Durable Managed Session journal/failover; model/reasoning at session create; no known breaking changes
Read details→
C
CopilotKit@CopilotKit·21h ago
🛠️ Tooling

CopilotKit 1.74.0: consume learned skills across SDK containers; Web Inspector targeted What's New + HUD telemetry

CopilotKit shipped v1.74.0 on 2026-09-25 (npm @copilotkit/react-core 1.74.0): SDKs can consume learned skills from multiple containers, expanding skill assets beyond a single runtime; Web Inspector gains targeted What's New notifications and launcher HUD impression/action tracking. Fixes: threads drawer no longer locks on unresolved entitlements, stop refreshing realtime credentials for sockets that never open, stop per-message state cloning / virtual-scroll tug-of-war in React, and a Memory policy that grants nothing no longer fails the run.

CopilotKit 1.74.0: consume learned skills across SDK containers; Web Inspector targeted What's New + HUD telemetry
⚡ Key Takeaways
  • •Shipped: v1.74.0 / @copilotkit/[email protected]
  • •Skills: consume learned skills from multiple containers across SDKs (#7384)
  • •Inspector: targeted What's New (#6956) + launcher HUD impression/action tracking (#7376)
  • •Stability: no realtime credential refresh for never-open sockets; empty-grant Memory policy no longer fails runs
  • •React: stop per-message state cloning and virtual-scroll tug-of-war on long threads
Read details→
ADSponsored
P
Pydantic@pydantic·22h ago
🛠️ Tooling

Pydantic AI 2.51.0: OpenAILiveModel for GPT-Live, context_window_used on realtime, tighter Gemini Live connect checks

Pydantic AI v2.51.0 (2026-09-25) adds OpenAILiveModel for OpenAI GPT-Live, exposes context_window_used on realtime sessions, tightens Gemini Live id/affective-dialog checks, and raises UserError for forced realtime tool_choice on OpenAI/Azure/xAI. PyPI pydantic-ai 2.51.0 is live.

Pydantic AI 2.51.0: OpenAILiveModel for GPT-Live, context_window_used on realtime, tighter Gemini Live connect checks
⚡ Key Takeaways
  • •Shipped: GitHub v2.51.0 = PyPI pydantic-ai 2.51.0; docs pydantic.dev/docs/ai
  • •Realtime: OpenAILiveModel for GPT-Live; context_window_used on sessions from GPT-Live ratio or response usage
  • •Compat: match dated gemini-3.8-live ids; reject google_affective_dialog at connect; RealtimeError on Gemini Live close codes
  • •Safety: UserError for forced realtime tool_choice on OpenAI/Azure OpenAI/xAI
  • •Perf: cache agent graph across Agent.run(); skip task group for single-child fan-out
Read details→
A
Anthropic@AnthropicAI·23h ago
🛠️ Tooling

Claude Code 2.1.283: availableModelsMatch/deniedModels controls, gateway prompt-id headers, /doctor prompt-audit

Claude Code v2.1.283 (2026-09-25) adds availableModelsMatch=exact and deniedModels for managed model allow/deny lists, opt-in x-claude-code-prompt-id gateway hint headers, /doctor prompt-audit for CLAUDE.md/skills/agents/commands, plus Bedrock Mantle upstream and load_test_mode on the Apps Gateway. npm @anthropic-ai/claude-code 2.1.283 is live.

Claude Code 2.1.283: availableModelsMatch/deniedModels controls, gateway prompt-id headers, /doctor prompt-audit
⚡ Key Takeaways
  • •Shipped: GitHub v2.1.283 = npm @anthropic-ai/[email protected]
  • •Governance: availableModelsMatch=exact pins exact model versions; deniedModels blocks models even if allowlisted
  • •Observability: CLAUDE_CODE_GATEWAY_HINT_HEADERS=1 adds x-claude-code-prompt-id; OTEL_LOG_TOOL_CONTENT=1 logs tool outputs on spans
  • •Audit: /doctor prompt-audit (alias /checkup) reviews CLAUDE.md, skills, agents, commands for older-model patterns
  • •Gateway: Bedrock Mantle upstream + opt-in load_test_mode for signed canned replies
Read details→
S
Strands Agents@strands-agents·1d ago
🛠️ Tooling

Strands Agents 1.57.1: bidi snapshot capture/restore, vended a2a_client tool, CLI Quickstart/Customize/Export + /tools

Strands Agents (harness-sdk) python/v1.57.1 (2026-09-25) adds bidi snapshot capture/restore, a vended a2a_client tool, Agent.shutdown() for scope cleanup, CLI Quickstart/Customize/Export replacing Q&A setup plus /tools, and MCP TS client 2.0. PyPI strands-agents 1.57.1 is live.

Strands Agents 1.57.1: bidi snapshot capture/restore, vended a2a_client tool, CLI Quickstart/Customize/Export + /tools
⚡ Key Takeaways
  • •Shipped: harness-sdk python/v1.57.1 = PyPI strands-agents 1.57.1; docs strandsagents.com/docs
  • •Bidi: snapshot capture/restore; explicit model IDs for providers; align response boundaries and record tool dispatch in history
  • •Interop: vended a2a_client tool; extract content from task.status.message when artifacts are absent
  • •Lifecycle: Agent.shutdown() for scope cleanup; MCP server migrates off removed mcp.server.fastmcp for mcp 2.x
  • •CLI: Quickstart/Customize/Export replace setup Q&A; add /tools; companion harness-cli/v0.1.4 same day
Read details→
G
Google ADK@google·1d ago
🛠️ Tooling

Google ADK Python 2.10.0: experimental skill lifecycle caps, MongoDB vector/hybrid search, eval duration/token/call metrics

Google ADK Python v2.10.0 (GitHub 2026-09-25; changelog 2026-09-24) adds experimental skill lifecycle controls (EPHEMERAL, active-skill caps, unload_skill via ADK_ENABLE_SKILL_LIFECYCLE=1), a MongoDB toolset with vector/hybrid search, eval efficiency metrics for duration/tokens/model calls, and better OpenAI reasoning param adaptation plus reasoning-token reporting. PyPI google-adk 2.10.0 is live.

Google ADK Python 2.10.0: experimental skill lifecycle caps, MongoDB vector/hybrid search, eval duration/token/call metrics
⚡ Key Takeaways
  • •Shipped: GitHub v2.10.0 + PyPI google-adk 2.10.0; docs at adk.dev/get-started
  • •Skills (experimental): enable ADK_ENABLE_SKILL_LIFECYCLE=1 for EPHEMERAL one-turn lifecycle, active-skill caps, and opt-in unload_skill
  • •Data: MongoDB toolset with vector and hybrid search inside agent flows
  • •Eval: duration efficiency metric plus token and model-call counts
  • •Models: OpenAI reasoning param adaptation and reasoning-token reporting; use effort instead of thinking_config on OpenAIResponsesLlm; adk create adds gemini-3.8-flash
Read details→
M
Mem0@mem0ai·1d ago
🛠️ Tooling

Mem0 Python SDK 2.2.1: add() stops fake ADD successes; Bedrock Claude reasoning + Turbopuffer filter/score fixes

Mem0 Python SDK v2.2.1 (2026-09-25, PyPI mem0ai 2.2.1) makes `Memory`/`AsyncMemory.add()` report only vector-store inserts that actually succeeded—raising `VectorStoreError` when none insert—restores context-manager close semantics, fixes Bedrock Anthropic Converse to read the first text-bearing content block (after optional reasoningContent), applies full Turbopuffer filter operators, and remaps euclidean(_squared)/S3 Vectors scores to `1/(1+distance)`. Same-day Node SDK ts-v3.3.1 stops injecting OpenAI default baseURL/model into non-OpenAI providers.

Mem0 Python SDK 2.2.1: add() stops fake ADD successes; Bedrock Claude reasoning + Turbopuffer filter/score fixes
⚡ Key Takeaways
  • •Shipped: Python v2.2.1 / PyPI mem0ai 2.2.1; Node ts-v3.3.1 same day
  • •Correctness: add() returns only inserted memories; raises VectorStoreError if none insert
  • •Bedrock Anthropic: read first text-bearing Converse block after optional reasoningContent
  • •Turbopuffer filters: full eq/ne/gt/gte/lt/lte/in/nin; unsupported ops raise ValueError
  • •Scores: euclidean_squared and S3 Vectors euclidean map via 1/(1+distance) to avoid negatives/collapse-to-zero
Read details→
P
Prefect / FastMCP@PrefectHQ·1d ago
🛠️ Tooling

FastMCP 4.0.10: task tools stay registered behind Search/CodeMode; cross-tool call_tool returns real results

Prefect shipped FastMCP v4.0.10 (2026-09-25, PyPI fastmcp 4.0.10, “Inside Job”): task-enabled tools remain registered with the task backend even when hidden by Search transforms or CodeMode, and calls via `ctx.fastmcp.call_tool()` (including the search `call_tool` proxy and CodeMode `execute`) now return real results instead of empty task receipts. Also tolerates stdio transport construction failures in `__del__`. Docs at gofastmcp.com.

FastMCP 4.0.10: task tools stay registered behind Search/CodeMode; cross-tool call_tool returns real results
⚡ Key Takeaways
  • •Shipped: FastMCP v4.0.10 / PyPI fastmcp 4.0.10
  • •Tasks: tools hidden by Search/CodeMode still register with the task backend (#5262)
  • •Calls: ctx.fastmcp.call_tool() / CodeMode execute return real results, not empty task receipts (#5275)
  • •Stdio: __del__ tolerates failed transport construction (#5256)
  • •Docs: gofastmcp.com + repo changelog updated for v4.0.10
Read details→
C
Comet@CometML·1d ago
🛠️ Tooling

Comet Opik 2.2.81: scored entities → annotation queues, gated free-form ClickHouse SQL user, 19 provider price tables

Comet Opik 2.2.81 (2026-09-25) routes scored entities into annotation queues, optionally provisions an extended free-form SQL ClickHouse user, registers 19 canonical providers so model prices load, and fixes Gemini usage, streamed Mistral reasoning, and slots=True dataclass encoding in the SDK.

Comet Opik 2.2.81: scored entities → annotation queues, gated free-form ClickHouse SQL user, 19 provider price tables
⚡ Key Takeaways
  • •Shipped 2.2.81 on 2026-09-25; pip install opik==2.2.81
  • •Evaluation loop: route scored entities into annotation queues
  • •Data plane: optional extended free-form SQL ClickHouse user behind a flag; drop spans FINAL on find_trace_stream / find_thread_by_id
  • •Cost: register 19 canonical providers so model prices load
  • •SDK: Gemini usage without candidates_token_count; keep streamed Mistral reasoning; encode slots dataclasses as dicts; opik.sh startup timeout 90s
Read details→
O
OpenHands@OpenHands·1d ago
🛠️ Tooling

OpenHands v1.24.0: workspace folder toggle, read-only shared automations, MCP OAuth kept on cloud save

OpenHands Agent Canvas v1.24.0 (2026-09-25) adds header workspace-folder toggles, read-only shared automation conversations on cloud, MCP OAuth credentials preserved on cloud saves (skip consent when tokens still work), OpenAI subscription caches scoped by backend, and Node.js ≥24.

OpenHands v1.24.0: workspace folder toggle, read-only shared automations, MCP OAuth kept on cloud save
⚡ Key Takeaways
  • •Shipped 2026-09-25: GitHub v1.24.0 + npm @openhands/[email protected]
  • •UX: toggle all workspace folders from Conversations header; shared automation conversations open read-only on cloud
  • •MCP: keep OAuth credentials on cloud saves; skip consent when tokens still work; render ACP tool-call blocks in chat cards
  • •Multi-backend: scope OpenAI subscription caches and settings cache by answering backend
  • •Install: npm i -g @openhands/agent-canvas requires Node ≥24; see docs.openhands.dev
Read details→
L
Langfuse@langfuse·1d ago
🛠️ Tooling

Langfuse v4.46.0: basic skill management (unstable REST), streamlined experiment side-by-side, 100-row tracing tables

Langfuse v4.46.0 (2026-09-25) ships basic skill management (PR #17803) via unstable REST (list/create/get/patch/delete skills, versions, file content), streamlined experiment side-by-side comparison, formatted JSON on dataset items, optional 100-row tracing page size, assistant respect for data-retention, and v3 dataset-run-items reads on replicas.

Langfuse v4.46.0: basic skill management (unstable REST), streamlined experiment side-by-side, 100-row tracing tables
⚡ Key Takeaways
  • •Shipped: Langfuse v4.46.0 (2026-09-25); skills PR #17803 merged same day
  • •Skills (unstable): /api/public/unstable/skills CRUD-ish endpoints for skill artifacts aimed at coding agents
  • •Experiments: streamlined side-by-side comparison; formatted JSON on dataset items
  • •Observability UX: optional 100-row tracing page size; plain-text header metrics; 300ms tooltips
  • •Compliance/perf: assistant respects data-retention; v3 dataset-run-items on read replica; docs langfuse.com/docs
Read details→
C
Cline@cline·1d ago
🛠️ Tooling

Cline Desktop v0.0.36: proxy-safe local hub, fit-to-screen UI, Bedrock GPT-6 without cross-region

Cline Desktop v0.0.36 (2026-09-25) skips system HTTP(S) proxies for local hub checks, fits oversized windows to small/scaled displays, loads plugin slash commands on workspace open, and runs Bedrock GPT-6/GPT-5.6 without cross-region inference (India via in. profiles).

Cline Desktop v0.0.36: proxy-safe local hub, fit-to-screen UI, Bedrock GPT-6 without cross-region
⚡ Key Takeaways
  • •Local hub checks bypass system HTTP(S) proxies (Clash/v2ray/corporate)
  • •Oversized windows shrink to fit and center on scaled 1080p displays
  • •Plugin slash commands load on workspace open with background retry
  • •Bedrock GPT-6/GPT-5.6 work without cross-region; India uses in. profiles
  • •Artifacts: deb/rpm/dmg/exe + signatures; Providers panel scrolls model lists
Read details→
D
DeepSeek@deepseek_ai·1d ago
🛠️ Tooling

Official DeepSeek Harness Desktop App Released: Bundles Cordis Plugin Core for Daemon-Free Local Agent Workflows

DeepSeek AI (@deepseek_ai) has officially launched the standalone desktop application for DeepSeek Harness (dsh) across macOS and Windows. Designed to eliminate the engineering overhead of managing CLI background daemons and local sandbox sockets, DeepSeek Harness Desktop integrates the web UI, the Cordis plugin execution engine, and isolated sandboxing into a native local client. It delivers out-of-the-box support for visual trajectory playback, arbitrary session branching, and multi-model sub-agent delegations.

Official DeepSeek Harness Desktop App Released: Bundles Cordis Plugin Core for Daemon-Free Local Agent Workflows
⚡ Key Takeaways
  • •Eliminates complex multi-process CLI terminal setups by packaging the Cordis agent engine into a standalone GUI.
  • •Integrates native visual trajectory playback, inspecting tool execution, file diffs, and chain-of-thought traces in milliseconds.
  • •Introduces deterministic session forking, enabling developers to branch and retry failed agent steps without invalidating global context.
  • •Supports heterogeneous sub-agent delegation, seamlessly piping sub-tasks to Claude Code, OpenCode, or local Ollama instances.
Read details→
H
Hugging Face@huggingface·1d ago
🛠️ Tooling

Hugging Face TRL 1.14.0: chunked log-probs for DPO/KTO/GRPO, DDP all-reduce fix; up to 31% lower peak VRAM on H100

TRL v1.14.0 (2026-09-25) removes `trl.losses` fused-loss forks: DPO/KTO/GRPO stream via `_ChunkedLogProbFunction` into each trainer's own loss, fixing formula drift and plain-DDP missing all-reduce. Ships an in-tree Triton logprob+entropy kernel (~14× vs selective_log_softmax+entropy_from_logits). On 1×H100, Qwen3-0.6B, bs=4, seq 512: chunked+precompute_ref_log_probs median step 0.1971s (−12.1% vs v1.13 fused) and 4.83GB peak (−31%). Removes six low-use experimental trainers; vLLM window moves to ≤0.30.0.

Hugging Face TRL 1.14.0: chunked log-probs for DPO/KTO/GRPO, DDP all-reduce fix; up to 31% lower peak VRAM on H100
⚡ Key Takeaways
  • •Shipped: GitHub v1.14.0 = PyPI trl 1.14.0; docs huggingface.co/docs/trl
  • •Architecture: remove trl.losses; DPO/KTO/GRPO via _ChunkedLogProbFunction into existing losses; use_liger_kernel still enables Liger model kernels
  • •Perf (official, 1×H100 / Qwen3-0.6B / bs=4 / seq512): 0.1971s (−12.1%) and 4.83GB (−31%) with chunked+precompute_ref_log_probs vs v1.13 fused
  • •Kernel: in-tree Triton logprob+entropy ~14× (0.89ms vs 12.8ms) on the stated H100 shape; default on CUDA/ROCm/XPU
  • •Breaking: six experimental trainers removed; plain-DDP KTO all-reduce fixed; vLLM window >=0.20.0,<=0.30.0
Read details→