# Audio

Published articles for Audio.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## A Speaker That Clips to City Poles and Leaves No Mark

DevFeed: [A Speaker That Clips to City Poles and Leaves No Mark](<https://devfeed.tech/articles/a-speaker-that-clips-to-city-poles-and-leaves-no-mark-34917.md>)

Original publisher: [Read original article](<https://www.yankodesign.com/2026/09/16/a-speaker-that-clips-to-city-poles-and-leaves-no-mark/>)

Author: Ida Torres

Published: 2026-09-16T22:30:15Z

Content type: article

Language: en

Sources: [Yanko Design](<https://devfeed.tech/sources/yanko-design.md>)

Topics: [systems](<https://devfeed.tech/topics/systems.md>)

Tags: [architecture](<https://devfeed.tech/tags/architecture.md>), [audio](<https://devfeed.tech/tags/audio.md>), [audio-technology-concept-designs-public-speaker](<https://devfeed.tech/tags/audio-technology-concept-designs-public-speaker.md>), [building](<https://devfeed.tech/tags/building.md>), [concept-designs](<https://devfeed.tech/tags/concept-designs.md>), [design](<https://devfeed.tech/tags/design.md>), [public](<https://devfeed.tech/tags/public.md>), [sound](<https://devfeed.tech/tags/sound.md>), [space](<https://devfeed.tech/tags/space.md>), [speaker](<https://devfeed.tech/tags/speaker.md>), [technology](<https://devfeed.tech/tags/technology.md>)

### AI overview

Mariami Kurtishvili's Orchid Sound System is a graduate thesis proposing stainless steel high-frequency horns that clip onto existing urban infrastructure without screws, drilling, or permanent marks. The project explores temporary communal sound in public spaces.

### Source excerpt

A Speaker That Clips to City Poles and Leaves No Mark Walk down almost any street in an American city without your headphones in, and you will notice how little of that walk actually belongs to...

## Kacey Musgraves' Quest Concert Shows How Good Immersive Music Can Be

DevFeed: [Kacey Musgraves' Quest Concert Shows How Good Immersive Music Can Be](<https://devfeed.tech/articles/kacey-musgraves-quest-concert-shows-how-good-immersive-music-can-be-35499.md>)

Original publisher: [Read original article](<https://www.uploadvr.com/kacey-musgraves-quest-concert-shows-how-good-immersive-music-can-be/>)

Author: Craig Storm

Published: 2026-09-16T21:26:59Z

Content type: opinion

Language: en

Sources: [UploadVR](<https://devfeed.tech/sources/uploadvr.md>)

Topics: [3D](<https://devfeed.tech/topics/3d.md>), [Meta](<https://devfeed.tech/topics/meta.md>)

Tags: [3d](<https://devfeed.tech/tags/3d.md>), [4k](<https://devfeed.tech/tags/4k.md>), [audio](<https://devfeed.tech/tags/audio.md>), [camera](<https://devfeed.tech/tags/camera.md>), [cameras](<https://devfeed.tech/tags/cameras.md>), [immersive-video](<https://devfeed.tech/tags/immersive-video.md>), [lighting](<https://devfeed.tech/tags/lighting.md>), [meta](<https://devfeed.tech/tags/meta.md>), [music](<https://devfeed.tech/tags/music.md>), [production](<https://devfeed.tech/tags/production.md>), [quality](<https://devfeed.tech/tags/quality.md>), [wi-fi](<https://devfeed.tech/tags/wi-fi.md>)

### AI overview

A review of Kacey Musgraves: Middle of Nowhere, a 49-minute made-for-VR concert filmed at Billy Bob's Texas for Quest 3. The reviewer finds its stereoscopic 3D presentation immersive and natural, with clear visuals over hotel Wi-Fi and strong audio.

### Source excerpt

We had to stop ourselves from applauding. Kacey Musgraves' new made-for-VR concert on Quest 3 shows just how good immersive music can be.

## Google introduces Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking

DevFeed: [Google introduces Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking](<https://devfeed.tech/articles/google-gemini-3-8-live-40886.md>)

Original publisher: [Read original article](<https://habr.com/ru/companies/selectel/news/1082998/>)

Author: techno\_mot (Selectel)

Published: 2026-09-16T13:40:43Z

Content type: news

Language: ru

Sources: [Tagir Valeev](<https://devfeed.tech/sources/tagir-valeev.md>)

Topics: [Google](<https://devfeed.tech/topics/google.md>), [speech-to-speech](<https://devfeed.tech/topics/speech-to-speech.md>), [Google AI](<https://devfeed.tech/topics/google-ai.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [API](<https://devfeed.tech/topics/api.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [gemini](<https://devfeed.tech/tags/gemini.md>), [google](<https://devfeed.tech/tags/google.md>), [llm](<https://devfeed.tech/tags/llm.md>), [realtime](<https://devfeed.tech/tags/realtime.md>), [selectel](<https://devfeed.tech/tags/selectel.md>), [speech-to-speech](<https://devfeed.tech/tags/speech-to-speech.md>), [synthid](<https://devfeed.tech/tags/synthid.md>), [tag-efc6fa45f1fe](<https://devfeed.tech/tags/tag-efc6fa45f1fe.md>), [voice](<https://devfeed.tech/tags/voice.md>)

### AI overview

Google introduced two native speech-to-speech models: Gemini 3.8 Live, focused on scale and cost, and Gemini 3.8 Live Extended Thinking, designed for multi-step tasks with reasoning during dialogue. The article discusses direct audio processing, asynchronous function calls, visual context, multilingual conversations, SynthID watermarking, pricing, and availability through the Gemini API and Google AI Studio.

### Source excerpt

15 сентября Google представила две нативные speech-to-speech модели, Gemini 3.8 Live и Gemini 3.8 Live Extended Thinking. Первая заточена под масштаб и цену, вторая -- под многошаговые задачи с рассуждением прямо в диалоге. Интереснее баллов то, что голосовые агенты наконец получили признаки продакшен-продукта. Асинхронные вызовы функций, предсказуемая цена, интеграции с тем, на чем такие системы реально собирают. Хочу напомнить, как это делалось раньше. Голосовой бот -- конвейер из трех сервисов. Распознавание речи, языковая модель и синтез. Задержки складываются, а вместе с текстом может потеряться все остальное -- интонация, скорость речи, эмоция, шум на фоне. Модель получает расшифровку и не знает, что собеседник злится. Проблемным было и перебивание. Пока распознавание не закрыло фразу, система вообще не понимает, что ее прервали. Нативная speech-to-speech модель работает с аудио напрямую и снимает оба ограничения разом. Читать далее

## sem-ai 0.4.0: smaller responses, safer access, and pre-flight checks

DevFeed: [sem-ai 0.4.0: smaller responses, safer access, and pre-flight checks](<https://devfeed.tech/articles/sem-ai-0-4-0-smaller-responses-safer-access-and-pre-flight-checks-30850.md>)

Original publisher: [Read original article](<https://semaphore.io/blog/sem-ai-0.4.0-smaller-responses,-safer-access,-and-pre-flight-checks>)

Author: Pete Miloravac

Published: 2026-09-16T12:01:04Z

Content type: release

Language: en

Sources: [Semaphore Engineering](<https://devfeed.tech/sources/semaphore-engineering.md>)

Topics: [AI-assisted coding](<https://devfeed.tech/topics/ai-assisted-coding.md>), [Claude Code](<https://devfeed.tech/topics/claude-code.md>), [Command-line interface](<https://devfeed.tech/topics/cli.md>), [Authorization](<https://devfeed.tech/topics/authorization.md>), [API](<https://devfeed.tech/topics/api.md>), [configuration](<https://devfeed.tech/topics/configuration.md>)

Tags: [ai-coding-agents](<https://devfeed.tech/tags/ai-coding-agents.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [claude-code](<https://devfeed.tech/tags/claude-code.md>), [command-line](<https://devfeed.tech/tags/command-line.md>), [configuration](<https://devfeed.tech/tags/configuration.md>), [permissions](<https://devfeed.tech/tags/permissions.md>), [pipelines](<https://devfeed.tech/tags/pipelines.md>), [product-news](<https://devfeed.tech/tags/product-news.md>), [release](<https://devfeed.tech/tags/release.md>), [video](<https://devfeed.tech/tags/video.md>)

### AI overview

Semaphore's sem-ai 0.4.0 release adds more compact workflow and pipeline responses, context switching across organizations and credentials, and support for managing pre-flight checks. The update is aimed at reducing context usage and enabling more targeted permissions for AI coding agents.

### Source excerpt

AI coding agents work best when they receive the right information without unnecessary noise. They also need clearly defined permissions and reliable safeguards for the changes they make. sem-ai 0.4.0 improves all three areas. The release introduces more compact pipeline responses, flexible context switching, and support for Semaphore pre-flight checks. Watch Nick demonstrate the new [...] The post sem-ai 0.4.0: smaller responses, safer access, and pre-flight checks appeared first on Semaphore.

## Gemini Live audio

DevFeed: [Gemini Live audio](<https://devfeed.tech/articles/gemini-live-audio-31180.md>)

Original publisher: [Read original article](<https://simonwillison.net/2026/Sep/15/gemini-live/>)

Author: Simon Willison

Published: 2026-09-15T22:47:07Z

Content type: tutorial

Language: en

Sources: [Simon Willison's Weblog](<https://devfeed.tech/sources/simon-willison-s-weblog.md>)

Topics: [speech-to-speech](<https://devfeed.tech/topics/speech-to-speech.md>), [WebSocket](<https://devfeed.tech/topics/websocket.md>), [Google AI](<https://devfeed.tech/topics/google-ai.md>), [browser](<https://devfeed.tech/topics/browser.md>), [Playback](<https://devfeed.tech/topics/playback.md>), [implementation](<https://devfeed.tech/topics/implementation.md>), [API](<https://devfeed.tech/topics/api.md>)

Tags: [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [browser](<https://devfeed.tech/tags/browser.md>), [gemini](<https://devfeed.tech/tags/gemini.md>), [gemini-196](<https://devfeed.tech/tags/gemini-196.md>), [generative-ai](<https://devfeed.tech/tags/generative-ai.md>), [generative-ai-1-982](<https://devfeed.tech/tags/generative-ai-1-982.md>), [google](<https://devfeed.tech/tags/google.md>), [google-416](<https://devfeed.tech/tags/google-416.md>), [llm-release](<https://devfeed.tech/tags/llm-release.md>), [llm-release-231](<https://devfeed.tech/tags/llm-release-231.md>), [llms](<https://devfeed.tech/tags/llms.md>), [llms-1-948](<https://devfeed.tech/tags/llms-1-948.md>), [model](<https://devfeed.tech/tags/model.md>), [models](<https://devfeed.tech/tags/models.md>), [openai](<https://devfeed.tech/tags/openai.md>), [playback](<https://devfeed.tech/tags/playback.md>), [release](<https://devfeed.tech/tags/release.md>), [speech-to-speech](<https://devfeed.tech/tags/speech-to-speech.md>), [speech-to-text](<https://devfeed.tech/tags/speech-to-text.md>), [speech-to-text-21](<https://devfeed.tech/tags/speech-to-text-21.md>), [tools](<https://devfeed.tech/tags/tools.md>), [tools-78](<https://devfeed.tech/tags/tools-78.md>), [ui](<https://devfeed.tech/tags/ui.md>), [voice](<https://devfeed.tech/tags/voice.md>), [websocket](<https://devfeed.tech/tags/websocket.md>), [websockets](<https://devfeed.tech/tags/websockets.md>), [websockets-21](<https://devfeed.tech/tags/websockets-21.md>)

### AI overview

The article describes a browser-based web UI for trying Google's Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking speech-to-speech models. The implementation supports model and voice selection, an optional system prompt, voice conversations, and interruption while the model is speaking. It uses no libraries, connecting to a WebSocket endpoint and using the Web Audio API for capture and playback.

### Source excerpt

Tool: Gemini Live audio Google released Gemini 3.8 Live and 3.8 Live Extended Thinking today - two new speech-to-speech models that are a similar shape to OpenAI's GPT-Live family. I pointed GPT-6 Astra Extra High at the documentation and had it build me this web UI for trying out the new models. You can select a model and voice preset, enter an optional system prompt and then start a voice conversation through your browser, including the ability to interrupt the model while it is talking. The implementation uses no libraries. It connects to the wss://generativelanguage.googleapis.com/ws/google.ai.generativelanguage.v1alpha.GenerativeService.BidiGenerateContent?key=... WebSocket endpoint and uses a Web Audio API AudioContext for both capture and playback. Here's the Gemini Live tutorial for getting started with that WebSockets API. Tags: google, tools, websockets, generative-ai, llms, gemini, llm-release, speech-to-text

## Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

DevFeed: [Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking](<https://devfeed.tech/articles/introducing-gemini-3-8-live-and-3-8-live-extended-thinking-26922.md>)

Original publisher: [Read original article](<https://deepmind.google/blog/introducing-gemini-3-8-live-and-3-8-live-extended-thinking/>)

Author: Tom Ouyang

Published: 2026-09-15T17:05:57Z

Content type: release

Language: en

Sources: [Google DeepMind News](<https://devfeed.tech/sources/google-deepmind-news.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [real-time](<https://devfeed.tech/topics/real-time.md>), [speech-to-speech](<https://devfeed.tech/topics/speech-to-speech.md>), [Google](<https://devfeed.tech/topics/google.md>), [API](<https://devfeed.tech/topics/api.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [cost](<https://devfeed.tech/tags/cost.md>), [efficiency](<https://devfeed.tech/tags/efficiency.md>), [gemini](<https://devfeed.tech/tags/gemini.md>), [google](<https://devfeed.tech/tags/google.md>), [none](<https://devfeed.tech/tags/none.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [speech-to-speech](<https://devfeed.tech/tags/speech-to-speech.md>), [tools](<https://devfeed.tech/tags/tools.md>)

### AI overview

Google introduces Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two models designed for near-real-time voice interaction and reasoning. The release describes visual grounding, multilingual conversation, background tool and API execution, and deeper reasoning for complex workflows.

### Source excerpt

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet, built for natural conversation.

## September's Windows 11 patch needs an emergency patch of its own

DevFeed: [September's Windows 11 patch needs an emergency patch of its own](<https://devfeed.tech/articles/september-s-windows-11-patch-needs-an-emergency-patch-of-its-own-26958.md>)

Original publisher: [Read original article](<https://www.theregister.com/on-prem/2026/09/15/septembers-windows-11-patch-needs-an-emergency-patch-of-its-own/5296567>)

Author: Richard Speed

Published: 2026-09-15T13:33:55Z

Content type: news

Language: en

Sources: [www.theregister.com - Articles](<https://devfeed.tech/sources/www-theregister-com-articles.md>)

Topics: [Windows 11](<https://devfeed.tech/topics/windows-11.md>), [Microsoft](<https://devfeed.tech/topics/microsoft.md>), [USB](<https://devfeed.tech/topics/usb.md>)

Tags: [audio](<https://devfeed.tech/tags/audio.md>), [hyper-v](<https://devfeed.tech/tags/hyper-v.md>), [microsoft](<https://devfeed.tech/tags/microsoft.md>), [on-prem](<https://devfeed.tech/tags/on-prem.md>), [patch-tuesday](<https://devfeed.tech/tags/patch-tuesday.md>), [rdp](<https://devfeed.tech/tags/rdp.md>), [usb](<https://devfeed.tech/tags/usb.md>), [windows-11](<https://devfeed.tech/tags/windows-11.md>)

### AI overview

Microsoft fixes Windows 11 problems affecting RDP and Hyper-V, but some USB audio devices remain silent.

### Source excerpt

Microsoft fixes RDP and Hyper-V fallout, but some USB audio stays silent

## Syitren R400 is a portable Bluetooth CD player with a modular, customizable design

DevFeed: [Syitren R400 is a portable Bluetooth CD player with a modular, customizable design](<https://devfeed.tech/articles/finally-a-cd-player-gen-z-actually-wants-to-own-26652.md>)

Original publisher: [Read original article](<https://www.yankodesign.com/2026/09/15/finally-a-cd-player-gen-z-actually-wants-to-own/>)

Author: Ida Torres

Published: 2026-09-15T13:20:31Z

Content type: opinion

Language: en

Sources: [Yanko Design](<https://devfeed.tech/sources/yanko-design.md>)

Topics: [Hardware](<https://devfeed.tech/topics/hardware.md>), [Bluetooth](<https://devfeed.tech/topics/bluetooth.md>)

Tags: [2026](<https://devfeed.tech/tags/2026.md>), [a-design-award-and-competition](<https://devfeed.tech/tags/a-design-award-and-competition.md>), [audio](<https://devfeed.tech/tags/audio.md>), [audio-music-technology-a-design-award-and-competition-cd-player-music-player](<https://devfeed.tech/tags/audio-music-technology-a-design-award-and-competition-cd-player-music-player.md>), [bluetooth](<https://devfeed.tech/tags/bluetooth.md>), [cd-player](<https://devfeed.tech/tags/cd-player.md>), [collection](<https://devfeed.tech/tags/collection.md>), [cost](<https://devfeed.tech/tags/cost.md>), [design](<https://devfeed.tech/tags/design.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [modular](<https://devfeed.tech/tags/modular.md>), [mount](<https://devfeed.tech/tags/mount.md>), [music](<https://devfeed.tech/tags/music.md>), [music-player](<https://devfeed.tech/tags/music-player.md>), [portable](<https://devfeed.tech/tags/portable.md>), [product](<https://devfeed.tech/tags/product.md>), [technology](<https://devfeed.tech/tags/technology.md>)

### AI overview

The article presents Syitren's R400 as an entry-level portable Bluetooth CD player designed by ORGDOT Design. Its detachable modular shell supports customization, a desk stand, and an optional wall-mount bracket, while dual-material construction and cable-management details aim to improve everyday use.

### Source excerpt

Finally, a CD Player Gen Z Actually Wants to Own It really warms my heart when I see Gen Z and Gen Alpha kids collecting physical media from their favorite artists. These are people who...

## Firefox 157 Beta Released With Latest Enhancements

DevFeed: [Firefox 157 Beta Released With Latest Enhancements](<https://devfeed.tech/articles/firefox-157-beta-released-with-latest-enhancements-26760.md>)

Original publisher: [Read original article](<https://www.phoronix.com/news/Firefox-157-Beta>)

Author: Michael Larabel

Published: 2026-09-15T12:31:03Z

Content type: release

Language: en

Sources: [Phoronix](<https://devfeed.tech/sources/phoronix.md>)

Topics: [Firefox](<https://devfeed.tech/topics/firefox.md>), [Mozilla](<https://devfeed.tech/topics/mozilla.md>), [CSS](<https://devfeed.tech/topics/css.md>), [Nova](<https://devfeed.tech/topics/nova.md>), [Playback](<https://devfeed.tech/topics/playback.md>), [HTML](<https://devfeed.tech/topics/html.md>), [ui](<https://devfeed.tech/topics/ui.md>)

Tags: [audio](<https://devfeed.tech/tags/audio.md>), [css](<https://devfeed.tech/tags/css.md>), [desktop-linux](<https://devfeed.tech/tags/desktop-linux.md>), [firefox](<https://devfeed.tech/tags/firefox.md>), [html](<https://devfeed.tech/tags/html.md>), [linux-benchmarking](<https://devfeed.tech/tags/linux-benchmarking.md>), [linux-hardware-benchmarks](<https://devfeed.tech/tags/linux-hardware-benchmarks.md>), [linux-hardware-reviews](<https://devfeed.tech/tags/linux-hardware-reviews.md>), [linux-how-to](<https://devfeed.tech/tags/linux-how-to.md>), [linux-performance](<https://devfeed.tech/tags/linux-performance.md>), [linux-server-benchmarks](<https://devfeed.tech/tags/linux-server-benchmarks.md>), [mozilla](<https://devfeed.tech/tags/mozilla.md>), [nova](<https://devfeed.tech/tags/nova.md>), [open-source-graphics](<https://devfeed.tech/tags/open-source-graphics.md>), [phoronix](<https://devfeed.tech/tags/phoronix.md>), [phoronix-test-suite](<https://devfeed.tech/tags/phoronix-test-suite.md>), [playback](<https://devfeed.tech/tags/playback.md>), [release](<https://devfeed.tech/tags/release.md>), [ubuntu-benchmarks](<https://devfeed.tech/tags/ubuntu-benchmarks.md>), [ubuntu-hardware](<https://devfeed.tech/tags/ubuntu-hardware.md>), [ui](<https://devfeed.tech/tags/ui.md>)

### AI overview

Firefox 157 Beta adds improved HDR video handling for 8-bit formats, better audio/video synchronization when playback speed changes, and fixes for the vertical tabs sidebar. It also adds CSS support for the @supports at-rule function and the "chain" value for the CSS overscroll-behavior property, alongside Nova UI design work and the return of compact mode.

### Source excerpt

The Firefox release train keeps on rolling. With Firefox 156 released, Firefox 157 is now in beta ahead of the planned release coming up at the end of September...

## Apple is Working on Two iPhone Game Controllers Under its Beats Brand

DevFeed: [Apple is Working on Two iPhone Game Controllers Under its Beats Brand](<https://devfeed.tech/articles/apple-is-working-on-two-iphone-game-controllers-under-its-beats-brand-10834.md>)

Original publisher: [Read original article](<https://www.yankodesign.com/2026/09/13/apple-is-working-on-two-iphone-game-controllers-under-its-beats-brand/>)

Author: Sarang Sheth

Published: 2026-09-13T20:45:46Z

Content type: article

Language: en

Sources: [Yanko Design](<https://devfeed.tech/sources/yanko-design.md>)

Topics: [iphone](<https://devfeed.tech/topics/iphone.md>), [Hardware](<https://devfeed.tech/topics/hardware.md>), [Steam Deck](<https://devfeed.tech/topics/steam-deck.md>), [macOS](<https://devfeed.tech/topics/macos.md>)

Tags: [apple](<https://devfeed.tech/tags/apple.md>), [audio](<https://devfeed.tech/tags/audio.md>), [beats](<https://devfeed.tech/tags/beats.md>), [devices](<https://devfeed.tech/tags/devices.md>), [gadgets](<https://devfeed.tech/tags/gadgets.md>), [gadgets-gaming-product-design-apple-beats-gaming-controllers](<https://devfeed.tech/tags/gadgets-gaming-product-design-apple-beats-gaming-controllers.md>), [gaming](<https://devfeed.tech/tags/gaming.md>), [gaming-controllers](<https://devfeed.tech/tags/gaming-controllers.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [iphone](<https://devfeed.tech/tags/iphone.md>), [product](<https://devfeed.tech/tags/product.md>), [product-design](<https://devfeed.tech/tags/product-design.md>)

### AI overview

Apple is reportedly developing two iPhone game controllers under its Beats brand. The unreleased T6502 is wired, while the Bluetooth-enabled T1057 adds motion sensors and a different haptic engine. Neither device has a confirmed price, release date, or launch status.

### Source excerpt

Apple is Working on Two iPhone Game Controllers Under its Beats Brand Apple and the gaming crowd have always had one of those relationships where you nod at each other across a party and never actually walk...

## The improbable music of the ZX Spectrum's one-bit speaker

DevFeed: [The improbable music of the ZX Spectrum's one-bit speaker](<https://devfeed.tech/articles/the-improbable-music-of-the-zx-spectrum-s-one-bit-speaker-9007.md>)

Original publisher: [Read original article](<https://www.theregister.com/offbeat/2026/09/13/the-improbable-music-of-the-zx-spectrums-one-bit-speaker/5295862>)

Author: Liam Proven

Published: 2026-09-13T13:00:00Z

Content type: article

Language: en

Sources: [www.theregister.com - Articles](<https://devfeed.tech/sources/www-theregister-com-articles.md>)

Topics: [Hardware](<https://devfeed.tech/topics/hardware.md>)

Tags: [audio](<https://devfeed.tech/tags/audio.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [offbeat](<https://devfeed.tech/tags/offbeat.md>), [retro-computing](<https://devfeed.tech/tags/retro-computing.md>), [speakers](<https://devfeed.tech/tags/speakers.md>), [zx-spectrum](<https://devfeed.tech/tags/zx-spectrum.md>)

### AI overview

Ingenious coders coaxed multichannel chiptunes from the ZX Spectrum's one-bit speaker.

### Source excerpt

Ingenious coders coaxed multichannel chiptunes from the humblest of sound hardware

## How to Evaluate Live & Voice Agents in ADK

DevFeed: [How to Evaluate Live & Voice Agents in ADK](<https://devfeed.tech/articles/how-to-evaluate-live-voice-agents-in-adk-4212.md>)

Original publisher: [Read original article](<https://developers.googleblog.com/how-to-evaluate-live-voice-agents-in-adk/>)

Author: Stephen Allen

Published: 2026-09-12T11:04:33.891311Z

Content type: tutorial

Language: en

Sources: [Google Developers Blog](<https://devfeed.tech/sources/google-developers-blog.md>)

Topics: [AI Bots](<https://devfeed.tech/topics/ai-bots.md>), [CI/CD](<https://devfeed.tech/topics/cicd.md>)

Tags: [agents](<https://devfeed.tech/tags/agents.md>), [audio](<https://devfeed.tech/tags/audio.md>), [ci-cd](<https://devfeed.tech/tags/ci-cd.md>), [cli](<https://devfeed.tech/tags/cli.md>), [evaluation](<https://devfeed.tech/tags/evaluation.md>), [gemini](<https://devfeed.tech/tags/gemini.md>), [how-to](<https://devfeed.tech/tags/how-to.md>), [json](<https://devfeed.tech/tags/json.md>), [llm](<https://devfeed.tech/tags/llm.md>), [production](<https://devfeed.tech/tags/production.md>), [testing](<https://devfeed.tech/tags/testing.md>), [tool](<https://devfeed.tech/tags/tool.md>), [tools](<https://devfeed.tech/tags/tools.md>), [transcripts](<https://devfeed.tech/tags/transcripts.md>), [voice](<https://devfeed.tech/tags/voice.md>), [workflows](<https://devfeed.tech/tags/workflows.md>)

### AI overview

The article explains how to evaluate live voice agents in ADK with simulated audio conversations, automated scoring, and recorded results. It covers scenario-based and fixed-conversation test cases, multi-agent workflows, and running evaluations in CI/CD.

### Source excerpt

Moving live voice agents from demo to production requires rigorous, automated testing to handle the unpredictability of real multi-turn conversations. ADK now provides native live evaluation, allowing developers to test graph-based agent workflows against LLM-driven simulated users that generate actual audio via Gemini TTS. By defining evaluation scenarios and natural-language rubrics, you can automatically score audio responses and tool executions, inspect the resulting transcripts in ADK Web, or run the CLI directly in your CI/CD pipeline.

## BeOS-Inspired Haiku Now Supports Changing Audio Outputs Live, Other Improvements

DevFeed: [BeOS-Inspired Haiku Now Supports Changing Audio Outputs Live, Other Improvements](<https://devfeed.tech/articles/beos-inspired-haiku-now-supports-changing-audio-outputs-live-other-improvements-12408.md>)

Original publisher: [Read original article](<https://www.phoronix.com/news/Haiku-OS-August-2026>)

Author: Michael Larabel

Published: 2026-09-11T00:55:29Z

Content type: news

Language: en

Sources: [Phoronix](<https://devfeed.tech/sources/phoronix.md>)

Topics: [Operating system](<https://devfeed.tech/topics/operating-system.md>), [USB](<https://devfeed.tech/topics/usb.md>), [cpu](<https://devfeed.tech/topics/cpu.md>)

Tags: [2026](<https://devfeed.tech/tags/2026.md>), [architectures](<https://devfeed.tech/tags/architectures.md>), [audio](<https://devfeed.tech/tags/audio.md>), [build-system](<https://devfeed.tech/tags/build-system.md>), [cpu](<https://devfeed.tech/tags/cpu.md>), [desktop-linux](<https://devfeed.tech/tags/desktop-linux.md>), [development](<https://devfeed.tech/tags/development.md>), [kernel](<https://devfeed.tech/tags/kernel.md>), [linux-benchmarking](<https://devfeed.tech/tags/linux-benchmarking.md>), [linux-hardware-benchmarks](<https://devfeed.tech/tags/linux-hardware-benchmarks.md>), [linux-hardware-reviews](<https://devfeed.tech/tags/linux-hardware-reviews.md>), [linux-how-to](<https://devfeed.tech/tags/linux-how-to.md>), [linux-performance](<https://devfeed.tech/tags/linux-performance.md>), [linux-server-benchmarks](<https://devfeed.tech/tags/linux-server-benchmarks.md>), [open-source-graphics](<https://devfeed.tech/tags/open-source-graphics.md>), [os](<https://devfeed.tech/tags/os.md>), [phoronix](<https://devfeed.tech/tags/phoronix.md>), [phoronix-test-suite](<https://devfeed.tech/tags/phoronix-test-suite.md>), [power-management](<https://devfeed.tech/tags/power-management.md>), [release](<https://devfeed.tech/tags/release.md>), [report](<https://devfeed.tech/tags/report.md>), [ubuntu-benchmarks](<https://devfeed.tech/tags/ubuntu-benchmarks.md>), [ubuntu-hardware](<https://devfeed.tech/tags/ubuntu-hardware.md>), [update](<https://devfeed.tech/tags/update.md>), [usb](<https://devfeed.tech/tags/usb.md>)

### AI overview

Haiku's August 2026 development report describes live audio-output switching, automatic stopping of inactive outputs, continued Bluetooth and USB audio work, power-management improvements, CPU feature-detection updates, kernel fixes, build-system updates, and progress on PowerPC and ARM64 support.

### Source excerpt

In addition to August bringing the long-awaited Haiku R1 Beta 6 release, there was also a lot of development progress on this BeOS-inspired operating system too during the course of the past month...

## OpenAI arms devs with AI conversation tool that can talk and listen at the same time

DevFeed: [OpenAI arms devs with AI conversation tool that can talk and listen at the same time](<https://devfeed.tech/articles/openai-arms-devs-with-ai-conversation-tool-that-can-talk-and-listen-at-the-same-time-8532.md>)

Original publisher: [Read original article](<https://www.theregister.com/ai-and-ml/2026/09/10/openai-arms-devs-with-ai-conversation-tool-that-can-talk-and-listen-at-the-same-time/5295708>)

Author: Thomas Claburn

Published: 2026-09-10T22:59:00Z

Content type: news

Language: en

Sources: [www.theregister.com - Articles](<https://devfeed.tech/sources/www-theregister-com-articles.md>)

Topics: [Machine learning](<https://devfeed.tech/topics/machine-learning.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-and-ml](<https://devfeed.tech/tags/ai-and-ml.md>), [ai-models](<https://devfeed.tech/tags/ai-models.md>), [api](<https://devfeed.tech/tags/api.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [audio](<https://devfeed.tech/tags/audio.md>), [openai](<https://devfeed.tech/tags/openai.md>), [tool](<https://devfeed.tech/tags/tool.md>), [voice-ai](<https://devfeed.tech/tags/voice-ai.md>)

### AI overview

OpenAI's GPT-Live-1 is presented as a conversation tool that lets developers speak with AI models more fluidly.

### Source excerpt

GPT-Live-1 makes speaking to AI models more fluid

## Video and image search in Amazon Bedrock Knowledge Base using Marengo 3.0

DevFeed: [Video and image search in Amazon Bedrock Knowledge Base using Marengo 3.0](<https://devfeed.tech/articles/video-and-image-search-in-amazon-bedrock-knowledge-base-using-marengo-3-0-4743.md>)

Original publisher: [Read original article](<https://aws.amazon.com/blogs/machine-learning/video-and-image-search-in-amazon-bedrock-knowledge-base-using-marengo-3-0/>)

Author: Eric Kim

Published: 2026-09-10T21:15:39Z

Content type: tutorial

Language: en

Sources: [Artificial Intelligence](<https://devfeed.tech/sources/artificial-intelligence.md>)

Topics: [AI search](<https://devfeed.tech/topics/ai-search.md>), [Retrieval Augmented Generation (RAG)](<https://devfeed.tech/topics/retrieval-augmented-generation-rag.md>), [Amazon S3](<https://devfeed.tech/topics/amazon-s3.md>), [AWS IAM](<https://devfeed.tech/topics/aws-iam.md>)

Tags: [amazon-bedrock](<https://devfeed.tech/tags/amazon-bedrock.md>), [amazon-bedrock-knowledge-bases](<https://devfeed.tech/tags/amazon-bedrock-knowledge-bases.md>), [announcements](<https://devfeed.tech/tags/announcements.md>), [audio](<https://devfeed.tech/tags/audio.md>), [embedding](<https://devfeed.tech/tags/embedding.md>), [how-to](<https://devfeed.tech/tags/how-to.md>), [images](<https://devfeed.tech/tags/images.md>), [intermediate-200](<https://devfeed.tech/tags/intermediate-200.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [rag](<https://devfeed.tech/tags/rag.md>), [s3](<https://devfeed.tech/tags/s3.md>), [search](<https://devfeed.tech/tags/search.md>), [video](<https://devfeed.tech/tags/video.md>)

### AI overview

A walkthrough for building an Amazon Bedrock Knowledge Base with TwelveLabs Marengo Embed 3.0 to perform natural-language semantic search across video, images, and audio.

### Source excerpt

TwelveLabs Marengo Embed 3.0 is now generally available as an embedding model in Amazon Bedrock Knowledge Bases, bringing fully managed natural language search to video, image, and audio content. This walkthrough shows how to build a knowledge base powered by Marengo 3.0 and run semantic queries against your media.

## Arduino UNO Media Carrier adds MIPI CSI/DSI and audio connectors to UNO Q and VENTUNO Q boards

DevFeed: [Arduino UNO Media Carrier adds MIPI CSI/DSI and audio connectors to UNO Q and VENTUNO Q boards](<https://devfeed.tech/articles/arduino-uno-media-carrier-adds-mipi-csi-dsi-and-audio-connectors-to-uno-q-and-ventuno-q-boards-14037.md>)

Original publisher: [Read original article](<https://www.cnx-software.com/2026/09/10/arduino-uno-media-carrier-adds-mipi-csi-dsi-and-audio-connectors-to-uno-q-and-ventuno-q-boards/>)

Author: Jean-Luc Aufranc (CNXSoft)

Published: 2026-09-10T00:00:32Z

Content type: news

Language: en

Sources: [CNX Software - Embedded Systems News](<https://devfeed.tech/sources/cnx-software-embedded-systems-news.md>)

Topics: [UNO Q](<https://devfeed.tech/topics/uno-q.md>), [Arduino](<https://devfeed.tech/topics/arduino.md>), [Raspberry Pi](<https://devfeed.tech/topics/raspberry-pi.md>), [Hardware](<https://devfeed.tech/topics/hardware.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>)

Tags: [arduino](<https://devfeed.tech/tags/arduino.md>), [audio](<https://devfeed.tech/tags/audio.md>), [camera](<https://devfeed.tech/tags/camera.md>), [connectors](<https://devfeed.tech/tags/connectors.md>), [display](<https://devfeed.tech/tags/display.md>), [embedded-systems](<https://devfeed.tech/tags/embedded-systems.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [linux](<https://devfeed.tech/tags/linux.md>), [mipi](<https://devfeed.tech/tags/mipi.md>), [mipi-csi](<https://devfeed.tech/tags/mipi-csi.md>), [news](<https://devfeed.tech/tags/news.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [single-board-computer](<https://devfeed.tech/tags/single-board-computer.md>), [specifications](<https://devfeed.tech/tags/specifications.md>), [touchscreen](<https://devfeed.tech/tags/touchscreen.md>), [uno-q](<https://devfeed.tech/tags/uno-q.md>), [ventuno-q](<https://devfeed.tech/tags/ventuno-q.md>), [waveshare](<https://devfeed.tech/tags/waveshare.md>)

### AI overview

The Arduino UNO Media Carrier adds accessible 22-pin MIPI DSI and CSI connectors, three 3.5 mm audio jacks, and RGB LEDs to Arduino UNO Q and VENTUNO Q boards. It is listed at $19.25 or EUR 19.89 including VAT.

### Source excerpt

Arduino UNO Media Carrier board adds three audio jacks, a MIPI DSI display connector, two MIPI CSI camera connectors, and a few RGB LEDs to the Arduino UNO Q and Arduino VENTUNO Q SBCs. Both Arduino "Q" boards already expose MIPI CSI/DSI and audio interfaces, but only through the JMEDIA and JMISC 60-pin headers, which aren't very convenient. The Arduino UNO Media Carrier board fixes that with 22-pin MIPI connectors and 3.5 audio jacks. Arduino UNO Media Carrier (ASX00083) specifications: Compatible SBCs - Arduino UNO Q and Arduino VENTUNO Q Display I/F - 22-pin MIPI DSI connector compatible with standard MIPI-DSI display modules (5-inch, 8-inch, 10.1-inch Waveshare touch display compatible) Camera I/F - 2x 22-pin MIPI-CSI connectors compatible with IMX219 camera modules like the Raspberry Pi Camera Module 2 Audio 3.5 mm (MIC-IN/Headphones Out) audio jack 3.5mm Line Out audio jack 3.5mm Ear Out audio jack Host interfaces 60-pin female [...] The post Arduino UNO Media Carrier adds MIPI CSI/DSI and audio connectors to UNO Q and VENTUNO Q boards appeared first on CNX Software - Embedded Systems News.

## Build more natural voice experiences with GPT-Live-1 in the API

DevFeed: [Build more natural voice experiences with GPT-Live-1 in the API](<https://devfeed.tech/articles/build-more-natural-voice-experiences-with-gpt-live-1-in-the-api-6496.md>)

Original publisher: [Read original article](<https://openai.com/index/introducing-gpt-live-1-in-the-api>)

Published: 2026-09-10T00:00:00Z

Content type: release

Language: en

Sources: [OpenAI News](<https://devfeed.tech/sources/openai-news.md>)

Topics: [voice ai](<https://devfeed.tech/topics/voice-ai.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>), [Latency](<https://devfeed.tech/topics/latency.md>), [asr](<https://devfeed.tech/topics/asr.md>)

Tags: [agents](<https://devfeed.tech/tags/agents.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [gpt](<https://devfeed.tech/tags/gpt.md>), [latency](<https://devfeed.tech/tags/latency.md>), [llm](<https://devfeed.tech/tags/llm.md>), [product](<https://devfeed.tech/tags/product.md>), [reasoning](<https://devfeed.tech/tags/reasoning.md>), [release](<https://devfeed.tech/tags/release.md>), [text-to-speech](<https://devfeed.tech/tags/text-to-speech.md>), [tool](<https://devfeed.tech/tags/tool.md>), [voice](<https://devfeed.tech/tags/voice.md>)

### AI overview

GPT-Live-1 is launching in the API for full-duplex voice applications. The release emphasizes interruption handling, customizable conversational behavior, delegated reasoning and tool calls, long-session reliability, and telephony support.

### Source excerpt

GPT-Live-1 brings natural, full-duplex voice conversations to the API, with stronger instruction following, custom voices, and telephony support.

## Build an AI Voice Assistant with Twilio Voice and Media Streams, OpenAI's GPT-Live API, and Node.js

DevFeed: [Build an AI Voice Assistant with Twilio Voice and Media Streams, OpenAI's GPT-Live API, and Node.js](<https://devfeed.tech/articles/build-an-ai-voice-assistant-with-twilio-voice-and-media-streams-openai-s-gpt-live-api-and-node-js-16093.md>)

Original publisher: [Read original article](<https://www.twilio.com/en-us/blog/developers/tutorials/integrations/voice-ai-assistant-openai-gpt-live-1-node>)

Author: Paul Kamp

Published: 2026-09-10T00:00:00Z

Content type: tutorial

Language: en

Sources: [Twilio Blog](<https://devfeed.tech/sources/twilio-blog.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Node.js](<https://devfeed.tech/topics/node-js.md>), [OpenAI](<https://devfeed.tech/topics/openai.md>), [Streams](<https://devfeed.tech/topics/streams.md>), [API](<https://devfeed.tech/topics/api.md>), [WebSocket](<https://devfeed.tech/topics/websocket.md>), [Fastify](<https://devfeed.tech/topics/fastify.md>), [API keys](<https://devfeed.tech/topics/api-keys.md>), [.env](<https://devfeed.tech/topics/dotenv.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [developer-insights](<https://devfeed.tech/tags/developer-insights.md>), [env-file-security](<https://devfeed.tech/tags/env-file-security.md>), [js](<https://devfeed.tech/tags/js.md>), [ngrok](<https://devfeed.tech/tags/ngrok.md>), [node-js](<https://devfeed.tech/tags/node-js.md>), [openai](<https://devfeed.tech/tags/openai.md>), [sign-up](<https://devfeed.tech/tags/sign-up.md>), [streams](<https://devfeed.tech/tags/streams.md>)

### AI overview

This tutorial explains how to build an AI voice assistant that answers phone calls using Twilio Programmable Voice and Media Streams, OpenAI's GPT-Live-1 API, and a Node.js server. It also covers web search and custom function tool calls.

### Source excerpt

Build an AI voice assistant that answers a phone call with Twilio Programmable Voice and Media Streams, powered by OpenAI's GPT-Live-1.

## LG TV flaws could let attackers listen in, even in standby mode

DevFeed: [LG TV flaws could let attackers listen in, even in standby mode](<https://devfeed.tech/articles/lg-tv-flaws-could-let-attackers-listen-in-even-in-standby-mode-8445.md>)

Original publisher: [Read original article](<https://www.malwarebytes.com/blog/privacy/2026/09/lg-tv-flaws-could-let-attackers-listen-in-even-in-standby-mode>)

Author: Pieter Arntz

Published: 2026-09-07T13:34:40Z

Content type: news

Language: en

Sources: [Malwarebytes](<https://devfeed.tech/sources/malwarebytes.md>)

Topics: [Vulnerabilities](<https://devfeed.tech/topics/vulnerabilities.md>), [Networks](<https://devfeed.tech/topics/networks.md>), [Embedded Software Dev](<https://devfeed.tech/topics/embedded-software-dev.md>)

Tags: [acr](<https://devfeed.tech/tags/acr.md>), [audio](<https://devfeed.tech/tags/audio.md>), [lg](<https://devfeed.tech/tags/lg.md>), [networks](<https://devfeed.tech/tags/networks.md>), [news](<https://devfeed.tech/tags/news.md>), [privacy](<https://devfeed.tech/tags/privacy.md>), [security](<https://devfeed.tech/tags/security.md>), [smart-tv](<https://devfeed.tech/tags/smart-tv.md>), [testing](<https://devfeed.tech/tags/testing.md>), [vulnerabilities](<https://devfeed.tech/tags/vulnerabilities.md>)

### AI overview

An investigation reported that LG smart TVs may scan local networks and collect viewing-related data, while undisclosed vulnerabilities could enable audio capture from a compromised TV.

### Source excerpt

Testing found that LG smart TVs can track viewing and scan home networks, while security flaws could let attackers record conversations.

## Coding Challenge #135 - Voice Dictation App

DevFeed: [Coding Challenge #135 - Voice Dictation App](<https://devfeed.tech/articles/coding-challenge-135-voice-dictation-app-29210.md>)

Original publisher: [Read original article](<https://codingchallenges.substack.com/p/coding-challenge-135-voice-dictation>)

Author: John Crickett

Published: 2026-09-05T08:01:16Z

Content type: tutorial

Language: en

Sources: [Coding Challenges](<https://devfeed.tech/sources/coding-challenges.md>)

Topics: [Code Challenge](<https://devfeed.tech/topics/code-challenge.md>), [Programming](<https://devfeed.tech/topics/programming.md>), [AI Development](<https://devfeed.tech/topics/ai-development.md>), [Accessibility](<https://devfeed.tech/topics/accessibility.md>), [Operating system](<https://devfeed.tech/topics/operating-system.md>)

Tags: [accessibility](<https://devfeed.tech/tags/accessibility.md>), [ai](<https://devfeed.tech/tags/ai.md>), [audio](<https://devfeed.tech/tags/audio.md>), [coding](<https://devfeed.tech/tags/coding.md>), [os](<https://devfeed.tech/tags/os.md>), [programming](<https://devfeed.tech/tags/programming.md>), [ui-automation](<https://devfeed.tech/tags/ui-automation.md>), [voice](<https://devfeed.tech/tags/voice.md>)

### AI overview

A Coding Challenge tutorial for building a private voice dictation app that runs entirely on the local computer. It covers audio capture, local speech recognition, text cleanup, spoken formatting commands, custom dictionaries, history, and inserting dictated text into the focused application without network requests.

### Source excerpt

This challenge is to build your own voice dictation app.

## Introducing agentic video understanding with Gemini

DevFeed: [Introducing agentic video understanding with Gemini](<https://devfeed.tech/articles/introducing-agentic-video-understanding-with-gemini-6192.md>)

Original publisher: [Read original article](<https://deepmind.google/blog/introducing-agentic-video-in-gemini/>)

Author: Rohan Doshi

Published: 2026-09-01T17:08:51Z

Content type: release

Language: en

Sources: [Google DeepMind News](<https://devfeed.tech/sources/google-deepmind-news.md>)

Topics: [Google AI](<https://devfeed.tech/topics/google-ai.md>), [Google](<https://devfeed.tech/topics/google.md>)

Tags: [agentic](<https://devfeed.tech/tags/agentic.md>), [ai](<https://devfeed.tech/tags/ai.md>), [analysis](<https://devfeed.tech/tags/analysis.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [benchmarks](<https://devfeed.tech/tags/benchmarks.md>), [cost](<https://devfeed.tech/tags/cost.md>), [developers](<https://devfeed.tech/tags/developers.md>), [enterprise](<https://devfeed.tech/tags/enterprise.md>), [feature](<https://devfeed.tech/tags/feature.md>), [flash](<https://devfeed.tech/tags/flash.md>), [gemini](<https://devfeed.tech/tags/gemini.md>), [google-ai](<https://devfeed.tech/tags/google-ai.md>), [models](<https://devfeed.tech/tags/models.md>), [none](<https://devfeed.tech/tags/none.md>), [performance](<https://devfeed.tech/tags/performance.md>), [reasoning](<https://devfeed.tech/tags/reasoning.md>), [retrieval](<https://devfeed.tech/tags/retrieval.md>), [tools](<https://devfeed.tech/tags/tools.md>), [transcripts](<https://devfeed.tech/tags/transcripts.md>), [video](<https://devfeed.tech/tags/video.md>)

### AI overview

Google launches agentic video understanding for Gemini Flash models. It uses native video tools to inspect relevant frames, audio, and transcripts, aiming to improve video-analysis accuracy while reducing token use and cost.

### Source excerpt

We're launching agentic video understanding across our latest Gemini models for improved accuracy and lower costs and token usage.

## MiniMax H3 and H3 Max are 50% off on AI Gateway

DevFeed: [MiniMax H3 and H3 Max are 50% off on AI Gateway](<https://devfeed.tech/articles/minimax-h3-and-h3-max-are-50-off-on-ai-gateway-1013.md>)

Original publisher: [Read original article](<https://vercel.com/changelog/minimax-h3-and-h3-max-are-50-off-on-ai-gateway>)

Author: Jerilyn Zheng

Published: 2026-08-30T00:00:00Z

Content type: release

Language: en

Sources: [Vercel News](<https://devfeed.tech/sources/vercel-news.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [API](<https://devfeed.tech/topics/api.md>), [API keys](<https://devfeed.tech/topics/api-keys.md>), [dashboards](<https://devfeed.tech/topics/dashboards.md>), [browser](<https://devfeed.tech/topics/browser.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [api](<https://devfeed.tech/tags/api.md>), [audio](<https://devfeed.tech/tags/audio.md>), [browser](<https://devfeed.tech/tags/browser.md>), [generate](<https://devfeed.tech/tags/generate.md>), [generation](<https://devfeed.tech/tags/generation.md>), [image](<https://devfeed.tech/tags/image.md>), [model](<https://devfeed.tech/tags/model.md>), [models](<https://devfeed.tech/tags/models.md>), [pricing](<https://devfeed.tech/tags/pricing.md>)

### AI overview

MiniMax H3 and H3 Max receive a 50% discount on Vercel AI Gateway from August 30 through September 13. H3 supports 2K video generation from text, images, video, and audio inputs, while H3 Max offers faster 480p and 768p rendering from text or a starting image. Existing model IDs remain unchanged, so no code changes are required.

### Source excerpt

MiniMax H3 and H3 Max are 50% off on AI Gateway from August 30 through September 13, in partnership with MiniMax. The discount covers requests billed through AI Gateway, at every duration and in every aspect ratio the model supports. H3 generates 2K video from a text prompt, a starting image, a pair of first and last frames, or reference images, video, and audio. H3 Max trades resolution for speed: it renders faster at 480p and 768p, and it takes a text prompt or a starting image. The model IDs (minimax/minimax-h3 and minimax/minimax-h3-max) are unchanged, so requests you already send pick up the discounted rate with no code change: Renders take minutes, so poll runs the generation as a background job and makes short status requests until it lands, rather than holding one long request open. See asynchronous generation for the webhook and start-and-status routes. Get started Create an API key in the AI Gateway section of your dashboard, or generate a clip in the browser first from the model playground. Current rates for every model are on the pricing page. You can view all video models available on AI Gateway. Read more

## Hedge 317: AI Tools and Coding

DevFeed: [Hedge 317: AI Tools and Coding](<https://devfeed.tech/articles/hedge-317-ai-tools-and-coding-10884.md>)

Original publisher: [Read original article](<https://rule11.tech/hedge-317/>)

Author: Russ

Published: 2026-08-28T17:39:02Z

Content type: article

Language: en

Sources: [rule 11 reader](<https://devfeed.tech/sources/rule-11-reader.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [coding](<https://devfeed.tech/topics/coding.md>), [Software](<https://devfeed.tech/topics/software.md>)

Tags: [agentic](<https://devfeed.tech/tags/agentic.md>), [ai](<https://devfeed.tech/tags/ai.md>), [ai-tools](<https://devfeed.tech/tags/ai-tools.md>), [audio](<https://devfeed.tech/tags/audio.md>), [coding](<https://devfeed.tech/tags/coding.md>), [developers](<https://devfeed.tech/tags/developers.md>), [hedge](<https://devfeed.tech/tags/hedge.md>)

### AI overview

A discussion of how effective AI and agentic AI tools are for building and testing software, drawing on project experience and lessons learned for developers.

### Source excerpt

How effective are AI tools--even agentic AI tools placed into a system--at building and testing software? Derick Winkworth and Donald Sharp Join Russ to discuss various projects they've worked on using AI tools, lessons they've learned, and how developers can make effective use of AI.

## The Open ASR Leaderboard Adds Its First Global South Language

DevFeed: [The Open ASR Leaderboard Adds Its First Global South Language](<https://devfeed.tech/articles/the-open-asr-leaderboard-adds-its-first-global-south-language-7411.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/open-asr-leaderboard-global-south>)

Author: Eric Bezzam; Shobhit Banga; Manas Dhir; Bhaskar Singh; Manmeet Kaur; Aaditya Pareek; Walecha; Sagar Jain; Hanuman Sidh; Vanshika Chhabra

Published: 2026-08-28T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [asr](<https://devfeed.tech/topics/asr.md>), [Benchmark](<https://devfeed.tech/topics/benchmark.md>), [Human-AI evaluation](<https://devfeed.tech/topics/human-ai-evaluation.md>)

Tags: [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [asr](<https://devfeed.tech/tags/asr.md>), [audio](<https://devfeed.tech/tags/audio.md>), [benchmark](<https://devfeed.tech/tags/benchmark.md>), [benchmarks](<https://devfeed.tech/tags/benchmarks.md>), [contributors](<https://devfeed.tech/tags/contributors.md>), [datasets](<https://devfeed.tech/tags/datasets.md>), [devices](<https://devfeed.tech/tags/devices.md>), [evaluation](<https://devfeed.tech/tags/evaluation.md>), [leaderboard](<https://devfeed.tech/tags/leaderboard.md>), [metrics](<https://devfeed.tech/tags/metrics.md>), [open](<https://devfeed.tech/tags/open.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [research](<https://devfeed.tech/tags/research.md>), [speech](<https://devfeed.tech/tags/speech.md>)

### AI overview

The Open ASR Leaderboard introduces Monsoon evaluation sets for Hindi in India, expanding coverage beyond European languages and testing how recognition performance varies across populations and conditions. The sets use public and private splits, speaker-disjoint data, detailed speaker attributes, and variation in geography, age, gender, vocabulary, devices, acoustic environments, speech type, speech rate, and transcript validity.

### Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.

[Next page](<https://devfeed.tech/tags/audio.md?cursor=WyIyMDI2LTA4LTI4VDAwOjAwOjAwKzAwOjAwIiwgImE4MmVkNTUyLTRhY2YtNDMyMy1hYjY0LTdhOTUxYjcyMzhkOCJd>)