# vision

Published articles for vision.

This is one page of public article previews, not the complete archive. Follow Next page to continue. Summaries are not the original full articles.

## Hackers Stole Flock's Camera Software, Revealing How the Company Tracks Cars and People

DevFeed: [Hackers Stole Flock's Camera Software, Revealing How the Company Tracks Cars and People](<https://devfeed.tech/articles/hackers-stole-flock-s-camera-software-revealing-how-the-company-tracks-cars-and-people-41550.md>)

Original publisher: [Read original article](<https://yro.slashdot.org/story/26/09/17/0517235/hackers-stole-flocks-camera-software-revealing-how-the-company-tracks-cars-and-people>)

Author: EditorDavid

Published: 2026-09-17T16:04:00Z

Content type: news

Language: en

Sources: [Slashdot](<https://devfeed.tech/sources/slashdot.md>)

Topics: [Computer vision](<https://devfeed.tech/topics/computer-vision.md>), [Encryption](<https://devfeed.tech/topics/encryption.md>), [Software](<https://devfeed.tech/topics/software.md>), [data](<https://devfeed.tech/topics/data.md>), [Cryptography](<https://devfeed.tech/topics/cryptography.md>)

Tags: [camera](<https://devfeed.tech/tags/camera.md>), [encryption](<https://devfeed.tech/tags/encryption.md>), [flock](<https://devfeed.tech/tags/flock.md>), [image](<https://devfeed.tech/tags/image.md>), [logs](<https://devfeed.tech/tags/logs.md>), [net-11](<https://devfeed.tech/tags/net-11.md>), [people](<https://devfeed.tech/tags/people.md>), [privacy](<https://devfeed.tech/tags/privacy.md>), [software](<https://devfeed.tech/tags/software.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

A joint investigation analyzed data copied from a Flock roadway camera after hackers breached the device. The recovered files showed that its on-device software detects people, vehicles, license plates, bicycles, and some graphics, while logs documented extensive image generation and vehicle activity.

### Source excerpt

"Hackers ripped down a Flock camera above a roadway, made a near-complete copy of the data stored inside it, and shared the files with 404 Media and WIRED," according to an article published on both sites. Though Flock has described its system as protected by on-device encryption, "The hackers were able to copy the camera's storage and recover an encryption key stored on the device, which unlocked videos of thousands of vehicle detections." The hackers shared the material with 404 Media and the transparency nonprofit Distributed Denial of Secrets, which shared the data with WIRED. 404 Media and WIRED then analyzed those files as part of a joint investigation... [T]he joint analysis of the recovered data shows that software running on the device explicitly detects people as well as vehicles, license plates, and bicycles. The camera can produce dozens of images of a single passing vehicle and, according to several weeks of recovered logs, generated more than a million images. Its computer-vision software also sometimes isolated bumper stickers and other graphics, including, in one case, an American flag patch on a motorcyclist's saddlebag... According to our analysis, the camera's logs recorded about 21 days of activity across several periods. During those windows, the device photographed roughly 50,200 vehicles and generated about 1.6 million images. On a typical day, it logged around 3,300 vehicles, with a high of 4,454... The software running on the camera explicitly detects people, something which is typically overlooked in discussions around Flock cameras. When it spots a person, it records where they appear in the image and how confident it is in the detection. It was a collective calling itself stegan0gram that breached the cameras, according to the interview they did with Wired and 404 Media. "Why just destroy them when we can reverse engineer them and find the secrets of those spying on us?" Read more of this story at Slashdot.

## New AI technique could make minimally invasive surgeries safer and more precise

DevFeed: [New AI technique could make minimally invasive surgeries safer and more precise](<https://devfeed.tech/articles/new-ai-technique-could-make-minimally-invasive-surgeries-safer-and-more-precise-37973.md>)

Original publisher: [Read original article](<https://news.mit.edu/2026/new-ai-technique-could-make-minimally-invasive-surgeries-safer-more-precise-0916>)

Author: Adam Zewe | MIT News

Published: 2026-09-16T15:00:00Z

Content type: news

Language: en

Sources: [MIT AI News](<https://devfeed.tech/sources/mit-ai-news.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [AI Research](<https://devfeed.tech/topics/ai-research.md>), [3D](<https://devfeed.tech/topics/3d.md>), [navigation](<https://devfeed.tech/topics/navigation.md>)

Tags: [3d](<https://devfeed.tech/tags/3d.md>), [ai](<https://devfeed.tech/tags/ai.md>), [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [computer-science-and-artificial-intelligence-laboratory-csail](<https://devfeed.tech/tags/computer-science-and-artificial-intelligence-laboratory-csail.md>), [computer-science-and-technology](<https://devfeed.tech/tags/computer-science-and-technology.md>), [computer-vision](<https://devfeed.tech/tags/computer-vision.md>), [electrical-engineering-and-computer-science-eecs](<https://devfeed.tech/tags/electrical-engineering-and-computer-science-eecs.md>), [health-care](<https://devfeed.tech/tags/health-care.md>), [images](<https://devfeed.tech/tags/images.md>), [imaging](<https://devfeed.tech/tags/imaging.md>), [jameel-clinic](<https://devfeed.tech/tags/jameel-clinic.md>), [machine-learning](<https://devfeed.tech/tags/machine-learning.md>), [medical-devices](<https://devfeed.tech/tags/medical-devices.md>), [medical-imaging](<https://devfeed.tech/tags/medical-imaging.md>), [minimally-invasive-surgery](<https://devfeed.tech/tags/minimally-invasive-surgery.md>), [mit-ibm-computing-research-lab](<https://devfeed.tech/tags/mit-ibm-computing-research-lab.md>), [mit-schwarzman-college-of-computing](<https://devfeed.tech/tags/mit-schwarzman-college-of-computing.md>), [model](<https://devfeed.tech/tags/model.md>), [national-institutes-of-health-nih](<https://devfeed.tech/tags/national-institutes-of-health-nih.md>), [navigation](<https://devfeed.tech/tags/navigation.md>), [paper](<https://devfeed.tech/tags/paper.md>), [polina-golland](<https://devfeed.tech/tags/polina-golland.md>), [precision](<https://devfeed.tech/tags/precision.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [research](<https://devfeed.tech/tags/research.md>), [school-of-engineering](<https://devfeed.tech/tags/school-of-engineering.md>), [vision](<https://devfeed.tech/tags/vision.md>), [vivek-gopalakrishnan](<https://devfeed.tech/tags/vivek-gopalakrishnan.md>)

### AI overview

MIT researchers and collaborators developed xvr, an AI method that adapts to individual patients and rapidly aligns intraoperative X-rays with preoperative 3D medical scans. The technique is intended to improve surgical navigation for minimally invasive procedures.

### Source excerpt

This patient-specific method, called xvr, helps doctors use X-rays for surgical navigation in fields such as orthopedics and neurosurgery.

## Designing Claude for ears, not eyes: An accessibility skill for blind and low-vision users

DevFeed: [Designing Claude for ears, not eyes: An accessibility skill for blind and low-vision users](<https://devfeed.tech/articles/designing-claude-for-ears-not-eyes-an-accessibility-skill-for-blind-and-low-vision-users-38847.md>)

Original publisher: [Read original article](<https://building.nubank.com/designing-claude-for-ears-not-eyes-an-accessibility-skill-for-blind-and-low-vision-users/>)

Author: Nubank Editorial

Published: 2026-09-16T14:09:16Z

Content type: article

Language: en

Sources: [Nubank](<https://devfeed.tech/sources/nubank.md>)

Topics: [Accessibility](<https://devfeed.tech/topics/accessibility.md>), [Claude](<https://devfeed.tech/topics/claude.md>), [User interface design](<https://devfeed.tech/topics/ui-design.md>), [screen](<https://devfeed.tech/topics/screen.md>), [Users](<https://devfeed.tech/topics/users.md>), [Development](<https://devfeed.tech/topics/development.md>)

Tags: [accessibility](<https://devfeed.tech/tags/accessibility.md>), [accessible](<https://devfeed.tech/tags/accessible.md>), [claude](<https://devfeed.tech/tags/claude.md>), [culture-values](<https://devfeed.tech/tags/culture-values.md>), [data-science-machine-learning](<https://devfeed.tech/tags/data-science-machine-learning.md>), [development-environment](<https://devfeed.tech/tags/development-environment.md>), [engineering](<https://devfeed.tech/tags/engineering.md>), [inclusive](<https://devfeed.tech/tags/inclusive.md>), [life-at-nu](<https://devfeed.tech/tags/life-at-nu.md>), [people-with-disabilities](<https://devfeed.tech/tags/people-with-disabilities.md>), [text](<https://devfeed.tech/tags/text.md>), [users](<https://devfeed.tech/tags/users.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Nubank describes BLIND-CLAUDE.md, an instruction set that adapts Claude responses for people who primarily use screen readers or read-aloud functionality. The approach treats auditory comprehension as an engineering constraint and changes behavioral instructions without modifying Claude or adding a native accessibility feature.

### Source excerpt

How a cognitively diverse IT team at Nubank, they coded screen reader-friendly rules into Claude's instructions, treating auditory comprehension as an engineering constraint The post Designing Claude for ears, not eyes: An accessibility skill for blind and low-vision users appeared first on Building Nubank.

## Designing Claude for ears, not eyes: An accessibility skill for blind and low-vision users

DevFeed: [Designing Claude for ears, not eyes: An accessibility skill for blind and low-vision users](<https://devfeed.tech/articles/designing-claude-for-ears-not-eyes-an-accessibility-skill-for-blind-and-low-vision-users-41432.md>)

Original publisher: [Read original article](<https://building.nu.com/designing-claude-for-ears-not-eyes-an-accessibility-skill-for-blind-and-low-vision-users/>)

Author: Nubank Editorial

Published: 2026-09-16T14:09:16Z

Content type: opinion

Language: en

Sources: [Nubank](<https://devfeed.tech/sources/nubank.md>)

Topics: [Accessibility](<https://devfeed.tech/topics/accessibility.md>), [Claude](<https://devfeed.tech/topics/claude.md>), [Users](<https://devfeed.tech/topics/users.md>), [screen](<https://devfeed.tech/topics/screen.md>)

Tags: [accessibility](<https://devfeed.tech/tags/accessibility.md>), [claude](<https://devfeed.tech/tags/claude.md>), [configuration](<https://devfeed.tech/tags/configuration.md>), [culture-values](<https://devfeed.tech/tags/culture-values.md>), [data-science-machine-learning](<https://devfeed.tech/tags/data-science-machine-learning.md>), [development-environment](<https://devfeed.tech/tags/development-environment.md>), [life-at-nu](<https://devfeed.tech/tags/life-at-nu.md>), [people-with-disabilities](<https://devfeed.tech/tags/people-with-disabilities.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Nubank describes BLIND-CLAUDE.md, an instruction set that adapts Claude responses for developers who primarily use screen readers or read-aloud tools. The approach addresses conventions such as Markdown tables, diffs, diagrams, long logs, and spatial references by adding behavioral instructions without modifying Claude or introducing a native accessibility feature.

### Source excerpt

How a cognitively diverse IT team at Nubank, they coded screen reader-friendly rules into Claude's instructions, treating auditory comprehension as an engineering constraint The post Designing Claude for ears, not eyes: An accessibility skill for blind and low-vision users appeared first on Building Nubank.

## CVITEK CV1842H-P-based edge AI camera module offers night vision and AI-ISP support (Crowdfunding)

DevFeed: [CVITEK CV1842H-P-based edge AI camera module offers night vision and AI-ISP support (Crowdfunding)](<https://devfeed.tech/articles/cvitek-cv1842h-p-based-edge-ai-camera-module-offers-night-vision-and-ai-isp-support-crowdfunding-27005.md>)

Original publisher: [Read original article](<https://www.cnx-software.com/2026/09/16/cvitek-cv1842h-p-based-edge-ai-camera-module-offers-night-vision-and-ai-isp-support/>)

Author: Debashis Das

Published: 2026-09-16T00:00:55Z

Content type: news

Language: en

Sources: [CNX Software - Embedded Systems News](<https://devfeed.tech/sources/cnx-software-embedded-systems-news.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Embedded Systems](<https://devfeed.tech/topics/embedded-systems.md>), [AI Inference](<https://devfeed.tech/topics/ai-inference.md>), [V](<https://devfeed.tech/topics/v.md>), [Arm](<https://devfeed.tech/topics/arm.md>), [cpu](<https://devfeed.tech/topics/cpu.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [Ethernet](<https://devfeed.tech/topics/ethernet.md>), [Linux](<https://devfeed.tech/topics/linux.md>), [Toit](<https://devfeed.tech/topics/toit.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [arm](<https://devfeed.tech/tags/arm.md>), [artificial-intelligence-ai](<https://devfeed.tech/tags/artificial-intelligence-ai.md>), [camera](<https://devfeed.tech/tags/camera.md>), [computer-vision](<https://devfeed.tech/tags/computer-vision.md>), [cpu](<https://devfeed.tech/tags/cpu.md>), [debug](<https://devfeed.tech/tags/debug.md>), [edge-ai](<https://devfeed.tech/tags/edge-ai.md>), [embedded](<https://devfeed.tech/tags/embedded.md>), [ethernet](<https://devfeed.tech/tags/ethernet.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [inference](<https://devfeed.tech/tags/inference.md>), [kickstarter](<https://devfeed.tech/tags/kickstarter.md>), [linux](<https://devfeed.tech/tags/linux.md>), [module](<https://devfeed.tech/tags/module.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [risc-v](<https://devfeed.tech/tags/risc-v.md>), [robotics](<https://devfeed.tech/tags/robotics.md>), [rt-thread](<https://devfeed.tech/tags/rt-thread.md>), [soc](<https://devfeed.tech/tags/soc.md>), [sophgo](<https://devfeed.tech/tags/sophgo.md>), [tinyml](<https://devfeed.tech/tags/tinyml.md>), [usb](<https://devfeed.tech/tags/usb.md>), [video](<https://devfeed.tech/tags/video.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

The AIMORELOGY Ovis is an open-source modular AI vision camera built around the CVITEK CV1842H-P SoC. It provides full-color 1080p night vision, 1.5 TOPS edge AI inference, USB and Ethernet connectivity, and a dual-OS environment using Linux and RT-Thread.

### Source excerpt

The AIMORELOGY Ovis is an open-source AI vision camera module built around the CVITEK CV1842H-P SoC, with full-color 1080p night vision and 1.5 TOPS edge AI inference in a compact modular design. It is designed for drones, robotics, security systems, smart cameras, and custom embedded vision products. The camera features a compact stacked design, with a 20 x 20 mm Core board that includes the CVITEK CV1842H-P SoC, 2 Gbit NAND flash, USB, and UART debug pads. A Sensor board with the SC235HAI image sensor connects on top and also adds Ethernet and UART interfaces. An optional CVBS board goes between the Sensor board and the Ovis Core board. AIMORELOGY Ovis specifications: Ovis Core Board SoC - CVITEK CV1842H-P CPU - 1x Arm Cortex-A53 core @ 1.1 GHz, 1x RISC-V C906 core @ 800 MHz NPU - 1.5 TOPS @ INT8 with BF16 support ISP - AI-ISP with real-time 1080p [...] The post CVITEK CV1842H-P-based edge AI camera module offers night vision and AI-ISP support (Crowdfunding) appeared first on CNX Software - Embedded Systems News.

## Arcturus Vision Camera Adds Color Passthrough & 3D Video Capture To Steam Frame

DevFeed: [Arcturus Vision Camera Adds Color Passthrough & 3D Video Capture To Steam Frame](<https://devfeed.tech/articles/arcturus-vision-camera-adds-color-passthrough-3d-video-capture-to-steam-frame-26798.md>)

Original publisher: [Read original article](<https://www.uploadvr.com/steam-frame-arcturus-vision-camera-color-passthrough/>)

Author: David Heaney

Published: 2026-09-15T06:59:00Z

Content type: news

Language: en

Sources: [UploadVR](<https://devfeed.tech/sources/uploadvr.md>)

Topics: [3D](<https://devfeed.tech/topics/3d.md>)

Tags: [3d](<https://devfeed.tech/tags/3d.md>), [camera](<https://devfeed.tech/tags/camera.md>), [capture](<https://devfeed.tech/tags/capture.md>), [experimental](<https://devfeed.tech/tags/experimental.md>), [headsets-tech](<https://devfeed.tech/tags/headsets-tech.md>), [rgb](<https://devfeed.tech/tags/rgb.md>), [video](<https://devfeed.tech/tags/video.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

The $149 Arcturus Vision Camera adds color passthrough to Valve's Steam Frame headset through its PCI-E expansion port. It also captures tracked stereoscopic 3D video, with options for 6-megapixel-per-eye HDR recording at 60fps or lower-resolution recording at 120fps.

### Source excerpt

The $149 Arcturus Vision Camera, weighing just 10 grams, adds 6 megapixel per eye color passthrough to Steam Frame and captures tracked 3D HDR video.

## Apple Laid Off Around 100 Staff Working On Vision Pro, But Don't Panic

DevFeed: [Apple Laid Off Around 100 Staff Working On Vision Pro, But Don't Panic](<https://devfeed.tech/articles/apple-laid-off-around-100-staff-working-on-vision-pro-but-don-t-panic-17469.md>)

Original publisher: [Read original article](<https://www.uploadvr.com/apple-vision-pro-layoffs-2026-gaming-immersive-video/>)

Author: David Heaney

Published: 2026-09-14T13:06:59Z

Content type: news

Language: en

Sources: [UploadVR](<https://devfeed.tech/sources/uploadvr.md>)

Topics: [visionOS](<https://devfeed.tech/topics/visionos.md>), [Security](<https://devfeed.tech/topics/security.md>), [Meta](<https://devfeed.tech/topics/meta.md>)

Tags: [apple](<https://devfeed.tech/tags/apple.md>), [article](<https://devfeed.tech/tags/article.md>), [company](<https://devfeed.tech/tags/company.md>), [cost](<https://devfeed.tech/tags/cost.md>), [ecosystem](<https://devfeed.tech/tags/ecosystem.md>), [gaming](<https://devfeed.tech/tags/gaming.md>), [immersive-video](<https://devfeed.tech/tags/immersive-video.md>), [industry-news](<https://devfeed.tech/tags/industry-news.md>), [meta](<https://devfeed.tech/tags/meta.md>), [open](<https://devfeed.tech/tags/open.md>), [report](<https://devfeed.tech/tags/report.md>), [security](<https://devfeed.tech/tags/security.md>), [series](<https://devfeed.tech/tags/series.md>), [third-party](<https://devfeed.tech/tags/third-party.md>), [video](<https://devfeed.tech/tags/video.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Apple laid off around 100 Vision Pro staff, largely shutting down its spatial gaming team and reducing its immersive video team. The cuts also affected security, spatial audio, Siri integration, and visionOS personnel, but Apple is not leaving the VR/MR market. The changes reflect limited gaming use and the high cost of producing immersive video. Apple may launch a slimmer, lighter successor headset around late 2028, although its continuation could depend on the wider headset market.

### Source excerpt

Apple laid off around 100 staff working on Vision Pro, including "largely shutting down" the spatial gaming team and downsizing the immersive video team, Bloomberg's Mark Gurman reported.

## NeoEyes NE302 - A tiny USB-C-powered WiFi 6 Edge AI Vision camera based on STM32N6 MCU

DevFeed: [NeoEyes NE302 - A tiny USB-C-powered WiFi 6 Edge AI Vision camera based on STM32N6 MCU](<https://devfeed.tech/articles/neoeyes-ne302-a-tiny-usb-c-powered-wifi-6-edge-ai-vision-camera-based-on-stm32n6-mcu-14041.md>)

Original publisher: [Read original article](<https://www.cnx-software.com/2026/09/12/neoeyes-ne302-a-tiny-usb-c-powered-wifi-6-edge-ai-vision-camera-based-on-stm32n6-mcu/>)

Author: Jean-Luc Aufranc (CNXSoft)

Published: 2026-09-12T02:19:36Z

Content type: news

Language: en

Sources: [CNX Software - Embedded Systems News](<https://devfeed.tech/sources/cnx-software-embedded-systems-news.md>)

Topics: [Embedded Systems](<https://devfeed.tech/topics/embedded-systems.md>), [Microcontroller](<https://devfeed.tech/topics/microcontroller.md>), [wifi 6](<https://devfeed.tech/topics/wifi-6.md>), [Hardware](<https://devfeed.tech/topics/hardware.md>), [object-detection](<https://devfeed.tech/topics/object-detection.md>)

Tags: [c-c-plus-plus](<https://devfeed.tech/tags/c-c-plus-plus.md>), [camera](<https://devfeed.tech/tags/camera.md>), [computer-vision](<https://devfeed.tech/tags/computer-vision.md>), [edge-ai](<https://devfeed.tech/tags/edge-ai.md>), [embedded-systems](<https://devfeed.tech/tags/embedded-systems.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [mcu](<https://devfeed.tech/tags/mcu.md>), [object-detection](<https://devfeed.tech/tags/object-detection.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [stm32](<https://devfeed.tech/tags/stm32.md>), [stmicro-stm32](<https://devfeed.tech/tags/stmicro-stm32.md>), [vision](<https://devfeed.tech/tags/vision.md>), [wifi-6](<https://devfeed.tech/tags/wifi-6.md>)

### AI overview

The article reports on CamThink's NeoEyes NE302, a compact USB-C-powered Wi-Fi 6 edge AI vision camera built around the STM32N6 MCU. It describes the device's reduced memory and feature set compared with the NE301, along with its camera, wireless, debugging, video-encoding, object-detection, networking, OTA, and security capabilities.

### Source excerpt

CamThink NeoEyes NE302 is a tiny WiFi 6 Edge AI camera based on an STM32N6 Arm Cortex-M55 MCU with Neural-ART NPU which the company says is "especially suitable for developers and makers". I initially thought about it as a smaller version of the CamThink NeoEyes NE301 based on the same STM32N6 MCU and 4MP OS04C10 camera sensor. But it's quite a different device. First, it doesn't have a battery, and the USB-C is only used for power. Memory and storage capacities have been reduced to 32 MB PSRAM and 64 MP SPI flash, and a range of built-in and optional features have been removed, including 4G LTE module (global or US), audio wafers, PIR motion connector, and a 16-pin GPIO header, as well as support for PoE power. CamThink NeoEyes NE302 specifications: Vision MCU STMicro STM32N6 MCU Core - Arm 32-bit Cortex-M55 CPU @ up to 800MHz with Arm Helium [...] The post NeoEyes NE302 - A tiny USB-C-powered WiFi 6 Edge AI Vision camera based on STM32N6 MCU appeared first on CNX Software - Embedded Systems News.

## iPhone Duo Seemingly Can't Capture Spatial Photos Or Video

DevFeed: [iPhone Duo Seemingly Can't Capture Spatial Photos Or Video](<https://devfeed.tech/articles/iphone-duo-seemingly-can-t-capture-spatial-photos-or-video-17283.md>)

Original publisher: [Read original article](<https://www.uploadvr.com/iphone-duo-seemingly-cant-capture-spatial-photos-or-video/>)

Author: Craig Storm

Published: 2026-09-11T21:15:11Z

Content type: news

Language: en

Sources: [UploadVR](<https://devfeed.tech/sources/uploadvr.md>)

Topics: [iphone](<https://devfeed.tech/topics/iphone.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Machine learning](<https://devfeed.tech/topics/machine-learning.md>)

Tags: [3d-media](<https://devfeed.tech/tags/3d-media.md>), [ai](<https://devfeed.tech/tags/ai.md>), [cameras](<https://devfeed.tech/tags/cameras.md>), [iphone](<https://devfeed.tech/tags/iphone.md>), [iphone-duo](<https://devfeed.tech/tags/iphone-duo.md>), [machine-learning](<https://devfeed.tech/tags/machine-learning.md>), [photos](<https://devfeed.tech/tags/photos.md>), [video](<https://devfeed.tech/tags/video.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Apple's $1,999 iPhone Duo appears not to support native spatial photo or spatial video capture, despite having cameras and processing hardware comparable to supported iPhone Pro models. Apple has not explained the omission, and the device was not independently tested.

### Source excerpt

Apple's $2000 iPhone Duo has two rear cameras, but seemingly can't capture spatial photos or video for viewing on Apple Vision Pro.

## DeepSeek V4.1 Flash now available on AI Gateway

DevFeed: [DeepSeek V4.1 Flash now available on AI Gateway](<https://devfeed.tech/articles/deepseek-v4-1-flash-now-available-on-ai-gateway-889.md>)

Original publisher: [Read original article](<https://vercel.com/changelog/deepseek-v4-1-flash-now-available-on-ai-gateway>)

Author: Jerilyn Zheng

Published: 2026-09-09T00:00:00Z

Content type: release

Language: en

Sources: [Vercel News](<https://devfeed.tech/sources/vercel-news.md>)

Topics: [deepseek](<https://devfeed.tech/topics/deepseek.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [API](<https://devfeed.tech/topics/api.md>), [Language models](<https://devfeed.tech/topics/language-models.md>), [AI-assisted coding](<https://devfeed.tech/topics/ai-assisted-coding.md>), [Claude Code](<https://devfeed.tech/topics/claude-code.md>), [codex](<https://devfeed.tech/topics/codex.md>), [cursor](<https://devfeed.tech/topics/cursor.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [api](<https://devfeed.tech/tags/api.md>), [caching](<https://devfeed.tech/tags/caching.md>), [codex](<https://devfeed.tech/tags/codex.md>), [context-window](<https://devfeed.tech/tags/context-window.md>), [cursor](<https://devfeed.tech/tags/cursor.md>), [deepseek](<https://devfeed.tech/tags/deepseek.md>), [responses](<https://devfeed.tech/tags/responses.md>), [tool](<https://devfeed.tech/tags/tool.md>), [vercel](<https://devfeed.tech/tags/vercel.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

DeepSeek V4.1 Flash is now available through Vercel's AI Gateway, offering native image understanding, a 1 million token context window, responses up to 384,000 tokens, reasoning, tool use, and prompt caching. Developers can use it in Claude Code, Codex, Cursor, and other coding agents with the model name deepseek/deepseek-v4.1-flash.

### Source excerpt

DeepSeek V4.1 Flash is now available on AI Gateway with native image understanding. V4.1 Flash has vision support and accepts text and images in the same request, so you can ask questions about screenshots, read charts, and extract information from visual content. The model has a 1 million token context window and supports responses up to 384,000 tokens, along with reasoning, tool use, and prompt caching. Its new architecture processes input and generates output with separate components, reducing the active computation needed for each stage. Use deepseek/deepseek-v4.1-flash as the model name: To use it in Claude Code, Codex, Cursor, and more, install the latest Vercel CLI and run setup: Then select deepseek/deepseek-v4.1-flash in the agent. See the coding agents guide for details. Try DeepSeek V4.1 Flash in the model playground. AI Gateway provides a unified API for calling models, tracking usage and cost, and configuring retries, failover, and performance optimizations for higher-than-provider uptime. It includes built-in custom reporting, Zero Data Retention support, budgets for API keys, routing rules, and more. AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests. You can view all language models available on AI Gateway. Read more

## Fovii Showcases Immersive Videos And Their Creators On Apple Vision Pro

DevFeed: [Fovii Showcases Immersive Videos And Their Creators On Apple Vision Pro](<https://devfeed.tech/articles/fovii-showcases-immersive-videos-and-their-creators-on-apple-vision-pro-17274.md>)

Original publisher: [Read original article](<https://www.uploadvr.com/fovii-showcases-immersive-videos-and-their-creators-on-apple-vision-pro/>)

Author: Laura Mingail

Published: 2026-09-08T14:00:00Z

Content type: article

Language: en

Sources: [UploadVR](<https://devfeed.tech/sources/uploadvr.md>)

Topics: [App](<https://devfeed.tech/topics/app.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [ui](<https://devfeed.tech/topics/ui.md>)

Tags: [apple](<https://devfeed.tech/tags/apple.md>), [business](<https://devfeed.tech/tags/business.md>), [creators](<https://devfeed.tech/tags/creators.md>), [immersive-video](<https://devfeed.tech/tags/immersive-video.md>), [launch](<https://devfeed.tech/tags/launch.md>), [subscriptions](<https://devfeed.tech/tags/subscriptions.md>), [ui](<https://devfeed.tech/tags/ui.md>), [video](<https://devfeed.tech/tags/video.md>), [video-streaming](<https://devfeed.tech/tags/video-streaming.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

The Fovii app is available on Apple Vision Pro as a platform for on-demand and live-streamed 180-degree immersive videos. It also highlights the creators, studios, and production work behind each project, with additional materials such as behind-the-scenes videos and spatial photos. The app is free to download, with monthly and annual subscriptions, individual rentals, and a limited launch selection.

### Source excerpt

The Fovii app on Apple Vision Pro offers a new platform to experience 180-degree immersive videos and discover the creators behind them.

## FunFitLand Brings Live Heart Rate Tracking Into VR With New Phone App

DevFeed: [FunFitLand Brings Live Heart Rate Tracking Into VR With New Phone App](<https://devfeed.tech/articles/funfitland-brings-live-heart-rate-tracking-into-vr-with-new-phone-app-17275.md>)

Original publisher: [Read original article](<https://www.uploadvr.com/funfitland-brings-live-heart-rate-tracking-into-vr-with-new-phone-app/>)

Author: Craig Storm

Published: 2026-09-08T02:00:00Z

Content type: news

Language: en

Sources: [UploadVR](<https://devfeed.tech/sources/uploadvr.md>)

Topics: [App](<https://devfeed.tech/topics/app.md>), [data](<https://devfeed.tech/topics/data.md>), [Android](<https://devfeed.tech/topics/android.md>), [iOS](<https://devfeed.tech/topics/ios.md>), [Mobile](<https://devfeed.tech/topics/mobile.md>), [Bluetooth](<https://devfeed.tech/topics/bluetooth.md>), [Meta](<https://devfeed.tech/topics/meta.md>), [Server](<https://devfeed.tech/topics/server.md>)

Tags: [android](<https://devfeed.tech/tags/android.md>), [app](<https://devfeed.tech/tags/app.md>), [apple](<https://devfeed.tech/tags/apple.md>), [bluetooth](<https://devfeed.tech/tags/bluetooth.md>), [compatibility](<https://devfeed.tech/tags/compatibility.md>), [fitness](<https://devfeed.tech/tags/fitness.md>), [meta](<https://devfeed.tech/tags/meta.md>), [phone-app](<https://devfeed.tech/tags/phone-app.md>), [servers](<https://devfeed.tech/tags/servers.md>), [smartphone](<https://devfeed.tech/tags/smartphone.md>), [testing](<https://devfeed.tech/tags/testing.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

FunFitLand's new iOS and Android companion app connects compatible wearables to VR workouts on Meta Quest and Apple Vision Pro, displaying live heart rate data and recording workout metrics.

### Source excerpt

FunFitLand's new smartphone app brings live heart rate data from compatible wearables directly into your VR workouts on Quest and Apple Vision Pro.

## Museas & Astronoma On Apple Vision Pro Support Curiosity & Discovery

DevFeed: [Museas & Astronoma On Apple Vision Pro Support Curiosity & Discovery](<https://devfeed.tech/articles/museas-astronoma-on-apple-vision-pro-support-curiosity-discovery-17292.md>)

Original publisher: [Read original article](<https://www.uploadvr.com/museas-astronoma-on-apple-vision-pro-support-curiosity-discovery/>)

Author: Laura Mingail

Published: 2026-09-07T23:28:57Z

Content type: article

Language: en

Sources: [UploadVR](<https://devfeed.tech/sources/uploadvr.md>)

Topics: [3D](<https://devfeed.tech/topics/3d.md>), [App](<https://devfeed.tech/topics/app.md>)

Tags: [3d](<https://devfeed.tech/tags/3d.md>), [apple](<https://devfeed.tech/tags/apple.md>), [apps](<https://devfeed.tech/tags/apps.md>), [art](<https://devfeed.tech/tags/art.md>), [design](<https://devfeed.tech/tags/design.md>), [discovery](<https://devfeed.tech/tags/discovery.md>), [education](<https://devfeed.tech/tags/education.md>), [exploration](<https://devfeed.tech/tags/exploration.md>), [interview](<https://devfeed.tech/tags/interview.md>), [learning](<https://devfeed.tech/tags/learning.md>), [science](<https://devfeed.tech/tags/science.md>), [technologies](<https://devfeed.tech/tags/technologies.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

An interview with Miguel Garcia Gonzalez examines how the Apple Vision Pro apps Museas and Astronoma use immersive design, real-world source material, and selective 3D presentation to encourage curiosity and exploration of art and science.

### Source excerpt

What makes exploring art and science in spatial apps so compelling? We look at Apple Vision Pro apps Museas and Astronoma to learn how to design for curiosity and discovery.

## Building a Real-Time 3D Face Mask with MediaPipe, Threlte and Three.js

DevFeed: [Building a Real-Time 3D Face Mask with MediaPipe, Threlte and Three.js](<https://devfeed.tech/articles/building-a-real-time-3d-face-mask-with-mediapipe-threlte-and-three-js-4342.md>)

Original publisher: [Read original article](<https://tympanus.net/codrops/2026/09/06/building-a-real-time-3d-face-mask-with-mediapipe-threlte-and-three-js/>)

Author: Marek Jóźwiak

Published: 2026-09-06T13:09:52Z

Content type: tutorial

Language: en

Sources: [Codrops](<https://devfeed.tech/sources/codrops.md>)

Topics: [Three.js webcam](<https://devfeed.tech/topics/three-js-webcam.md>)

Tags: [3d](<https://devfeed.tech/tags/3d.md>), [3d-face-mask](<https://devfeed.tech/tags/3d-face-mask.md>), [canonical-face-model](<https://devfeed.tech/tags/canonical-face-model.md>), [creative-coding](<https://devfeed.tech/tags/creative-coding.md>), [face-landmarks](<https://devfeed.tech/tags/face-landmarks.md>), [face-mesh](<https://devfeed.tech/tags/face-mesh.md>), [face-tracking](<https://devfeed.tech/tags/face-tracking.md>), [facial-landmarks](<https://devfeed.tech/tags/facial-landmarks.md>), [generative-graphics](<https://devfeed.tech/tags/generative-graphics.md>), [gpu](<https://devfeed.tech/tags/gpu.md>), [inference](<https://devfeed.tech/tags/inference.md>), [interactive-3d](<https://devfeed.tech/tags/interactive-3d.md>), [javascript-3d](<https://devfeed.tech/tags/javascript-3d.md>), [mediapipe](<https://devfeed.tech/tags/mediapipe.md>), [mediapipe-face-landmarker](<https://devfeed.tech/tags/mediapipe-face-landmarker.md>), [mesh](<https://devfeed.tech/tags/mesh.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [real-time-3d](<https://devfeed.tech/tags/real-time-3d.md>), [real-time-face-tracking](<https://devfeed.tech/tags/real-time-face-tracking.md>), [svelte](<https://devfeed.tech/tags/svelte.md>), [three-js](<https://devfeed.tech/tags/three-js.md>), [three-js-face-mask](<https://devfeed.tech/tags/three-js-face-mask.md>), [three-js-webcam](<https://devfeed.tech/tags/three-js-webcam.md>), [threlte](<https://devfeed.tech/tags/threlte.md>), [threlte-three-js](<https://devfeed.tech/tags/threlte-three-js.md>), [tutorial](<https://devfeed.tech/tags/tutorial.md>), [tutorials](<https://devfeed.tech/tags/tutorials.md>), [video](<https://devfeed.tech/tags/video.md>), [vision](<https://devfeed.tech/tags/vision.md>), [wasm](<https://devfeed.tech/tags/wasm.md>), [webcam](<https://devfeed.tech/tags/webcam.md>), [webcam-effects](<https://devfeed.tech/tags/webcam-effects.md>), [webcam-face-tracking](<https://devfeed.tech/tags/webcam-face-tracking.md>), [webgl](<https://devfeed.tech/tags/webgl.md>), [webgl-face-tracking](<https://devfeed.tech/tags/webgl-face-tracking.md>)

### AI overview

Tutorial on building a real-time textured 3D face mask by mapping MediaPipe face landmarks onto a Three.js mesh with Threlte.

### Source excerpt

Learn how MediaPipe's face landmarks, Google's canonical face model, and Three.js come together to create a real-time textured 3D face mask.

## Strengthening Camera Support in Zephyr for Advanced Vision Applications

DevFeed: [Strengthening Camera Support in Zephyr for Advanced Vision Applications](<https://devfeed.tech/articles/strengthening-camera-support-in-zephyr-for-advanced-vision-applications-13980.md>)

Original publisher: [Read original article](<https://www.zephyrproject.org/strengthening-camera-support-in-zephyr-for-advanced-vision-applications/>)

Author: Zephyr Project

Published: 2026-09-04T21:12:17Z

Content type: article

Language: en

Sources: [Zephyr Project](<https://devfeed.tech/sources/zephyr-project.md>)

Topics: [Embedded Systems](<https://devfeed.tech/topics/embedded-systems.md>), [Embedded Software Dev](<https://devfeed.tech/topics/embedded-software-dev.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [Framework](<https://devfeed.tech/topics/framework.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [object-detection](<https://devfeed.tech/topics/object-detection.md>), [real-time](<https://devfeed.tech/topics/real-time.md>), [Streaming](<https://devfeed.tech/topics/streaming.md>), [Robotics](<https://devfeed.tech/topics/robotics.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [blog](<https://devfeed.tech/tags/blog.md>), [cameras](<https://devfeed.tech/tags/cameras.md>), [development](<https://devfeed.tech/tags/development.md>), [embedded](<https://devfeed.tech/tags/embedded.md>), [embedded-systems](<https://devfeed.tech/tags/embedded-systems.md>), [events](<https://devfeed.tech/tags/events.md>), [india](<https://devfeed.tech/tags/india.md>), [industry-conference](<https://devfeed.tech/tags/industry-conference.md>), [object-detection](<https://devfeed.tech/tags/object-detection.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [open-source-summit](<https://devfeed.tech/tags/open-source-summit.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [robotics](<https://devfeed.tech/tags/robotics.md>), [streaming](<https://devfeed.tech/tags/streaming.md>), [vision](<https://devfeed.tech/tags/vision.md>), [zephyr](<https://devfeed.tech/tags/zephyr.md>), [zero-copy](<https://devfeed.tech/tags/zero-copy.md>)

### AI overview

This event recap examines proposed changes to Zephyr's camera and driver architecture for AI-driven vision workloads. The proposals include attaching metadata and inference results to individual video buffers and adding per-buffer callbacks to improve buffer ownership, reduce CPU wakeups, and support more manageable camera pipelines. The changes remain under exploration through prototypes and community discussions.

### Source excerpt

The Zephyr community came together at Open Source Summit India 2026 in Mumbai to share knowledge and explore developments, tooling, and real-world applications across embedded systems. In this second post event blog, we recap two lightning talks from the Zephyr track focused on camera support.

## Strengthening Camera Support in Zephyr for Advanced Vision Applications

DevFeed: [Strengthening Camera Support in Zephyr for Advanced Vision Applications](<https://devfeed.tech/articles/strengthening-camera-support-in-zephyr-for-advanced-vision-applications-38675.md>)

Original publisher: [Read original article](<https://zephyrproject.org/strengthening-camera-support-in-zephyr-for-advanced-vision-applications/>)

Author: Zephyr Project

Published: 2026-09-04T21:12:17Z

Content type: article

Language: en

Sources: [Zephyr Project](<https://devfeed.tech/sources/zephyr-project-2.md>)

Topics: [Embedded Software Dev](<https://devfeed.tech/topics/embedded-software-dev.md>), [Embedded Systems](<https://devfeed.tech/topics/embedded-systems.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>), [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [systems](<https://devfeed.tech/topics/systems.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [blog](<https://devfeed.tech/tags/blog.md>), [cameras](<https://devfeed.tech/tags/cameras.md>), [embedded](<https://devfeed.tech/tags/embedded.md>), [embedded-systems](<https://devfeed.tech/tags/embedded-systems.md>), [events](<https://devfeed.tech/tags/events.md>), [industry-conference](<https://devfeed.tech/tags/industry-conference.md>), [object-detection](<https://devfeed.tech/tags/object-detection.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [vision](<https://devfeed.tech/tags/vision.md>), [zephyr](<https://devfeed.tech/tags/zephyr.md>)

### AI overview

This event recap examines proposed changes to Zephyr's camera and driver architecture for AI vision workloads. It discusses metadata attached to video buffers, per-buffer callbacks, zero-copy sharing, bounded latency, backpressure, and streaming coordination. The proposals remain under exploration through prototypes and community discussions.

### Source excerpt

The Zephyr community came together at Open Source Summit India 2026 in Mumbai to share knowledge and explore developments, tooling, and real-world applications across embedded systems. In this second post event blog, we recap two lightning talks from the Zephyr track focused on camera support.

## ArduinoCore-Zephyr 1.0.0 adds support for Arduino VENTUNO Q, updates to Zephyr 4.4.1, and more

DevFeed: [ArduinoCore-Zephyr 1.0.0 adds support for Arduino VENTUNO Q, updates to Zephyr 4.4.1, and more](<https://devfeed.tech/articles/arduinocore-zephyr-1-0-0-adds-support-for-arduino-ventuno-q-updates-to-zephyr-4-4-1-and-more-14025.md>)

Original publisher: [Read original article](<https://www.cnx-software.com/2026/09/04/arduinocore-zephyr-1-0-0-adds-support-for-arduino-ventuno-q-updates-to-zephyr-4-4-1-and-more/>)

Author: Jean-Luc Aufranc (CNXSoft)

Published: 2026-09-04T10:27:38Z

Content type: release

Language: en

Sources: [CNX Software - Embedded Systems News](<https://devfeed.tech/sources/cnx-software-embedded-systems-news.md>)

Topics: [Arduino](<https://devfeed.tech/topics/arduino.md>), [Zephyr RTOS](<https://devfeed.tech/topics/zephyr-rtos.md>), [Embedded Systems](<https://devfeed.tech/topics/embedded-systems.md>), [Embedded Software Dev](<https://devfeed.tech/topics/embedded-software-dev.md>), [Hardware](<https://devfeed.tech/topics/hardware.md>), [Microcontroller](<https://devfeed.tech/topics/microcontroller.md>), [webcam](<https://devfeed.tech/topics/webcam.md>), [ci](<https://devfeed.tech/topics/ci.md>)

Tags: [arduino](<https://devfeed.tech/tags/arduino.md>), [bug-fixes](<https://devfeed.tech/tags/bug-fixes.md>), [camera](<https://devfeed.tech/tags/camera.md>), [ci](<https://devfeed.tech/tags/ci.md>), [development-board](<https://devfeed.tech/tags/development-board.md>), [embedded-systems](<https://devfeed.tech/tags/embedded-systems.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [mcu](<https://devfeed.tech/tags/mcu.md>), [real-time](<https://devfeed.tech/tags/real-time.md>), [release](<https://devfeed.tech/tags/release.md>), [software-management](<https://devfeed.tech/tags/software-management.md>), [ventuno-q](<https://devfeed.tech/tags/ventuno-q.md>), [vision](<https://devfeed.tech/tags/vision.md>), [zephyr](<https://devfeed.tech/tags/zephyr.md>), [zephyr-os](<https://devfeed.tech/tags/zephyr-os.md>)

### AI overview

Arduino has released ArduinoCore-Zephyr 1.0.0 with support for the VENTUNO Q SBC and Nicla Vision camera, an update to Zephyr 4.4.1, and bug-fix and CI improvements. The article also lists supported Arduino platforms and installation steps.

### Source excerpt

Arduino has just announced the release of ArduinoCore-Zephyr version 1.0.0, with new hardware support, including the VENTUNO Q SBC, an updated Zephyr 4.4.1 base, support for the Nicla Vision camera, and various fixes and improvements. Arduino first announced its plans to switch from Arm Mbed to Zephyr RTOS in July 2024, before releasing the first beta of Arduino Core for Zephyr in December of the same year. ArduinoCore-Zephyr 1.0.0 is not the first stable version, since that was version 0.90.0 released on July 29, 2026. The new release builds on that, so let's see what's new. ArduinoCore-Zephyr 1.0.0 highlights Support for Arduino VENTUNO Q SBC, specifically the STM32H5F5 real-time MCU on the board Update to Zephyr 4.4.1 with the latest upstream improvements, security fixes, and driver support Added Nicla Vision camera support to capture and process image data directly through the Zephyr core Various bug fixes and CI improvements Currently [...] The post ArduinoCore-Zephyr 1.0.0 adds support for Arduino VENTUNO Q, updates to Zephyr 4.4.1, and more appeared first on CNX Software - Embedded Systems News.

## NeoMME: an efficient Multimodal-native and Multilingual Encoder

DevFeed: [NeoMME: an efficient Multimodal-native and Multilingual Encoder](<https://devfeed.tech/articles/neomme-an-efficient-multimodal-native-and-multilingual-encoder-7011.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/Hcompany/neomme>)

Author: Tony Wu; Aurélien Lac

Published: 2026-09-03T13:13:48Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [vlm](<https://devfeed.tech/topics/vlm.md>), [Fine-tuning](<https://devfeed.tech/topics/fine-tuning.md>), [Large Language Model](<https://devfeed.tech/topics/llm.md>), [GPU](<https://devfeed.tech/topics/gpu.md>)

Tags: [apache](<https://devfeed.tech/tags/apache.md>), [architecture](<https://devfeed.tech/tags/architecture.md>), [diffusion](<https://devfeed.tech/tags/diffusion.md>), [embeddings](<https://devfeed.tech/tags/embeddings.md>), [fine-tuning](<https://devfeed.tech/tags/fine-tuning.md>), [gpu](<https://devfeed.tech/tags/gpu.md>), [hugging-face](<https://devfeed.tech/tags/hugging-face.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [quantization](<https://devfeed.tech/tags/quantization.md>), [retrieval](<https://devfeed.tech/tags/retrieval.md>), [training](<https://devfeed.tech/tags/training.md>), [transformers](<https://devfeed.tech/tags/transformers.md>), [vector](<https://devfeed.tech/tags/vector.md>), [vision](<https://devfeed.tech/tags/vision.md>), [vlm](<https://devfeed.tech/tags/vlm.md>)

### AI overview

NeoMME is a family of multilingual multimodal encoders trained from scratch with a masked discrete-diffusion objective. It uses one bidirectional Transformer for text tokens and image patches, and is fine-tuned for visual document retrieval with dense and late-interaction embeddings.

### Source excerpt

We introduce NeoMME, a family of 260M and 800M multilingual multimodal encoders. Unlike many generative visual language models, NeoMME does not use a separate pretrained vision tower or a causal language model. A single bidirectional Transformer processes both text tokens and raw image patches, and we train the entire model from scratch with a masked discrete-diffusion objective. We fine-tuned NeoMME for visual document retrieval using ColPali's page-image approach.

## Playco cut manual fixes 50% prototyping games with GPT-6 Astra

DevFeed: [Playco cut manual fixes 50% prototyping games with GPT-6 Astra](<https://devfeed.tech/articles/playco-cut-manual-fixes-50-prototyping-games-with-gpt-6-astra-6609.md>)

Original publisher: [Read original article](<https://openai.com/index/playco-game-prototyping-with-astra>)

Published: 2026-09-03T12:00:00Z

Content type: article

Language: en

Sources: [OpenAI News](<https://devfeed.tech/sources/openai-news.md>)

Topics: [Game engine](<https://devfeed.tech/topics/game-engine.md>), [ide](<https://devfeed.tech/topics/ide.md>), [Code](<https://devfeed.tech/topics/code.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-models](<https://devfeed.tech/tags/ai-models.md>), [bugs](<https://devfeed.tech/tags/bugs.md>), [game-development](<https://devfeed.tech/tags/game-development.md>), [gpt](<https://devfeed.tech/tags/gpt.md>), [ide](<https://devfeed.tech/tags/ide.md>), [performance](<https://devfeed.tech/tags/performance.md>), [prototypes](<https://devfeed.tech/tags/prototypes.md>), [prototyping](<https://devfeed.tech/tags/prototyping.md>), [reasoning](<https://devfeed.tech/tags/reasoning.md>), [startup](<https://devfeed.tech/tags/startup.md>), [ui](<https://devfeed.tech/tags/ui.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Playco reports that GPT-6 Astra helped it create three themed game prototypes from a shared grey-box foundation, with 50% fewer manual fixes than its previous model. The company uses the model in Playbot, an AI-powered IDE connected to game engines for editing, testing, validation, and bug finding.

### Source excerpt

Using GPT-6 Astra, Playco built three themed game prototypes from one grey box foundation and reported 50% fewer manual fixes than with the previous model.

## What Is the Raspberry Pi AI Kit? And What Replaced It?

DevFeed: [What Is the Raspberry Pi AI Kit? And What Replaced It?](<https://devfeed.tech/articles/what-is-the-raspberry-pi-ai-kit-and-what-replaced-it-10821.md>)

Original publisher: [Read original article](<https://raspberrytips.com/what-is-raspberry-pi-ai-kit/>)

Author: Dhairya Parikh

Published: 2026-09-03T01:38:42Z

Content type: tutorial

Language: en

Sources: [RaspberryTips](<https://devfeed.tech/sources/raspberrytips.md>)

Topics: [Hardware](<https://devfeed.tech/topics/hardware.md>), [Machine learning](<https://devfeed.tech/topics/machine-learning.md>), [Person Detection](<https://devfeed.tech/topics/person-detection.md>), [object-detection](<https://devfeed.tech/topics/object-detection.md>), [Linux](<https://devfeed.tech/topics/linux.md>), [Generative AI](<https://devfeed.tech/topics/generative-ai.md>)

Tags: [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [chatgpt](<https://devfeed.tech/tags/chatgpt.md>), [guide](<https://devfeed.tech/tags/guide.md>), [hardware](<https://devfeed.tech/tags/hardware.md>), [how-to-tutorials](<https://devfeed.tech/tags/how-to-tutorials.md>), [linux](<https://devfeed.tech/tags/linux.md>), [models](<https://devfeed.tech/tags/models.md>), [object-detection](<https://devfeed.tech/tags/object-detection.md>), [raspberry-pi](<https://devfeed.tech/tags/raspberry-pi.md>), [raspberry-pi-5](<https://devfeed.tech/tags/raspberry-pi-5.md>), [tutorial](<https://devfeed.tech/tags/tutorial.md>), [update](<https://devfeed.tech/tags/update.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

A tutorial explaining what the Raspberry Pi AI Kit was, how it enabled AI and computer vision projects on the Raspberry Pi 5, why it was discontinued, and what replaced it. The article also highlights its affordability, Hailo collaboration, software examples, and current availability.

### Source excerpt

When I first wrote about the Raspberry Pi AI Kit, it was one of the easiest ways to add AI acceleration to a Raspberry Pi 5. Well, things moved fast: Raspberry Pi has since discontinued it and introduced newer options instead. The official Raspberry Pi AI Kit made it much easier to run AI applications...

## PereStruct: Modular Pipeline and Dataset for Parsing Historical Newspapers

DevFeed: [PereStruct: Modular Pipeline and Dataset for Parsing Historical Newspapers](<https://devfeed.tech/articles/vlm-perestruct-24889.md>)

Original publisher: [Read original article](<https://habr.com/ru/companies/yandex/articles/1076770/>)

Author: makSShan (Яндекс, Yandex Cloud & Yandex Infrastructure)

Published: 2026-09-01T07:05:21Z

Content type: tutorial

Language: ru

Sources: [Яндекс - Как мы делаем Яндекс / Статьи](<https://devfeed.tech/sources/source.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [яндекс](<https://devfeed.tech/topics/tag-4004cf5948d3.md>), [vlm](<https://devfeed.tech/topics/vlm.md>)

Tags: [ai](<https://devfeed.tech/tags/ai.md>), [ai-studio](<https://devfeed.tech/tags/ai-studio.md>), [bleu](<https://devfeed.tech/tags/bleu.md>), [computervision](<https://devfeed.tech/tags/computervision.md>), [ocr](<https://devfeed.tech/tags/ocr.md>), [perestruct](<https://devfeed.tech/tags/perestruct.md>), [rouge](<https://devfeed.tech/tags/rouge.md>), [tag-4004cf5948d3](<https://devfeed.tech/tags/tag-4004cf5948d3.md>), [tag-65c8d6d9e736](<https://devfeed.tech/tags/tag-65c8d6d9e736.md>), [tag-9bf5e01ce62e](<https://devfeed.tech/tags/tag-9bf5e01ce62e.md>), [tag-d27a0708d400](<https://devfeed.tech/tags/tag-d27a0708d400.md>), [vision](<https://devfeed.tech/tags/vision.md>), [vlm](<https://devfeed.tech/tags/vlm.md>), [yandex-ai-studio](<https://devfeed.tech/tags/yandex-ai-studio.md>), [yolo](<https://devfeed.tech/tags/yolo.md>)

### AI overview

The article presents PereStruct, a modular pipeline for reconstructing articles from historical newspaper scans. It combines YOLO-based layout detection, Yandex Vision OCR, Yandex AI Studio models for error correction, and a semantic model for assembling article blocks; the authors also publish code, an annotated dataset, and a benchmark.

### Source excerpt

Попробуйте открыть скан советской газеты и прочитать одну статью от начала до конца. Человек быстро замечает крупный заголовок, продолжение в соседней колонке и подпись под фотографией. Для алгоритма перед ним -- это выцветшая страница с десятками тесно расположенных прямоугольников, нестандартными шрифтами и неоднозначным порядком чтения. Даже если OCR правильно распознаёт почти все слова, на выходе ещё не получится документ. Нужно понять, какие фрагменты относятся к одной статье, где её начало, в каком порядке соединить блоки и что не следует включать в основной текст. Мы разработали PereStruct -- модульный пайплайн для разбора исторических газет. Он объединяет детектор вёрстки на базе YOLO, Yandex Vision OCR, коррекцию ошибок с помощью моделей Yandex AI Studio и отдельную модель семантической сборки статей. Вместе с кодом мы публикуем размеченный датасет и бенчмарк, чтобы другие команды могли изучать подход и ставить эксперименты на исторических документах. Читать далее

## Qwen 3.8 Max 0902 now available on AI Gateway

DevFeed: [Qwen 3.8 Max 0902 now available on AI Gateway](<https://devfeed.tech/articles/qwen-3-8-max-0902-now-available-on-ai-gateway-1064.md>)

Original publisher: [Read original article](<https://vercel.com/changelog/qwen-3-8-max-0902-now-available-on-ai-gateway>)

Author: Jerilyn Zheng

Published: 2026-09-01T00:00:00Z

Content type: release

Language: en

Sources: [Vercel News](<https://devfeed.tech/sources/vercel-news.md>)

Topics: [qwen](<https://devfeed.tech/topics/qwen.md>), [Language models](<https://devfeed.tech/topics/language-models.md>), [AI-assisted coding](<https://devfeed.tech/topics/ai-assisted-coding.md>)

Tags: [ai-gateway](<https://devfeed.tech/tags/ai-gateway.md>), [coding-agents](<https://devfeed.tech/tags/coding-agents.md>), [models](<https://devfeed.tech/tags/models.md>), [qwen](<https://devfeed.tech/tags/qwen.md>), [release](<https://devfeed.tech/tags/release.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Qwen 3.8 Max 0902, a new Alibaba snapshot, is now available through AI Gateway. The release focuses on coding for larger projects, unsupervised long-horizon tasks, agent runs, and improved vision handling for charts and dense documents. It can be selected directly, pinned by its dated model ID, or reached through a rewrite routing rule. It is also supported in several coding-agent integrations and the model playground.

### Source excerpt

Qwen 3.8 Max 0902 from Alibaba is now available on AI Gateway. This is a new snapshot of Qwen 3.8 Max, with the gains concentrated in coding on larger projects, long-horizon work that runs without supervision, and agent runs. Vision handling is more accurate on charts and dense documents. To use Qwen 3.8 Max 0902, set model to alibaba/qwen3.8-max-0902: The dated ID pins this snapshot, so a later release will not change what your requests run against. To move existing traffic onto it without a code change, add a rewrite routing rule. The gateway substitutes the destination transparently, so an application that still asks for alibaba/qwen3.8-max runs on the new snapshot: To use it in a coding agent, see the coding agents guide, then run vercel ai-gateway coding-agents setup to connect agents like Claude Code, Codex, OpenCode, Cursor, Pi, and more and select alibaba/qwen3.8-max-0902. Try Qwen3.8-Max-0902 in the model playground. You can view all language models available on AI Gateway. Read more

## CodeScanner.scan(): Barcode Scanning Without Rebuilding the Camera Pipeline

DevFeed: [CodeScanner.scan(): Barcode Scanning Without Rebuilding the Camera Pipeline](<https://devfeed.tech/articles/codescanner-scan-barcode-scanning-without-rebuilding-the-camera-pipeline-19236.md>)

Original publisher: [Read original article](<https://www.codenameone.com/blog/camera-vision-scanners/>)

Author: Shai Almog

Published: 2026-08-26T00:00:00Z

Content type: release

Language: en

Sources: [CodeName One](<https://devfeed.tech/sources/codename-one.md>)

Topics: [Barcode](<https://devfeed.tech/topics/barcode.md>), [QR Code](<https://devfeed.tech/topics/qrcode.md>), [webcam](<https://devfeed.tech/topics/webcam.md>), [Code](<https://devfeed.tech/topics/code.md>)

Tags: [camera](<https://devfeed.tech/tags/camera.md>), [code](<https://devfeed.tech/tags/code.md>), [component](<https://devfeed.tech/tags/component.md>), [on-device](<https://devfeed.tech/tags/on-device.md>), [pipeline](<https://devfeed.tech/tags/pipeline.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Codename One adds CodeScanner and VisionCameraView, higher-level APIs for full-screen barcode scanning and embedded live vision analysis. The update also introduces typed results, coordinate helpers, image bridging, and analyzer-specific native dependency selection.

### Source excerpt

CodeScanner and VisionCameraView put full-screen scanning and embedded live analysis above Codename One's on-device vision APIs, with typed results and simulator scripting.

## Luce: Relightable Gaussians for 3D Asset Generation

DevFeed: [Luce: Relightable Gaussians for 3D Asset Generation](<https://devfeed.tech/articles/luce-relightable-gaussians-for-3d-asset-generation-6733.md>)

Original publisher: [Read original article](<https://machinelearning.apple.com/research/relightable-gaussians-3d-generation>)

Published: 2026-08-26T00:00:00Z

Content type: article

Language: en

Sources: [Apple Machine Learning Research](<https://devfeed.tech/sources/apple-machine-learning-research.md>)

Topics: [Artificial Intelligence](<https://devfeed.tech/topics/ai.md>), [Benchmark](<https://devfeed.tech/topics/benchmark.md>)

Tags: [3d](<https://devfeed.tech/tags/3d.md>), [ai](<https://devfeed.tech/tags/ai.md>), [benchmark](<https://devfeed.tech/tags/benchmark.md>), [computer-vision](<https://devfeed.tech/tags/computer-vision.md>), [generation](<https://devfeed.tech/tags/generation.md>), [images](<https://devfeed.tech/tags/images.md>), [mesh](<https://devfeed.tech/tags/mesh.md>), [models](<https://devfeed.tech/tags/models.md>), [multimodal](<https://devfeed.tech/tags/multimodal.md>), [techniques](<https://devfeed.tech/tags/techniques.md>), [vision](<https://devfeed.tech/tags/vision.md>)

### AI overview

Luce is a multimodal 3D representation for generating relightable assets from a single image. It combines geometry with physically based materials in a voxelized Gaussian cloud, compresses them into a material-aware latent space, and generates relightable PBR Gaussians and optional textured meshes. On Toys4K, it reports a 28% FID improvement over the strongest baseline and improves alignment on an AI-generated image benchmark.

### Source excerpt

High-fidelity image-to-3D generation requires a 3D representation that captures both geometry and appearance. To support relighting and integration into standard rendering pipelines, the representation should include physically based rendering (PBR) modalities such as albedo, metallic-roughness, and surface normals. We propose Luce, a 3D representation that unifies geometry and PBR materials within a voxelized multimodal Gaussian cloud, using dedicated Gaussian primitives for each modality. A variational autoencoder compresses this representation into a unified material-aware latent space. A...

[Next page](<https://devfeed.tech/tags/vision.md?cursor=WyIyMDI2LTA4LTI2VDAwOjAwOjAwKzAwOjAwIiwgIjE5MTFiMjFiLWEyYWQtNGE2NC1iYjM4LTgwZDM4OWQ2ZDcyYyJd>)