PromptAI

Vision — AI news & research

Image, video and multimodal AI — diffusion models, generation, segmentation and visual understanding. Updated continuously from 20+ curated sources.

AgentsVisionRoboticsPolicyResearchToolsLLMs
hn · Vision

Show HN: Kedge – Full-stack cloud with forkable VM snapshots and global SQLite

I'm building Kedge, a globally distributed platform for stateful serverless apps. Here's how you make a simple static site: `echo '# Hello world!' | ssh kedge.dev' I helped build Fly.io for 4 years a…

hn · Vision

Robotics development made dead simple (open source)

Hey everyone, My team and I have been working hard on this project: https://peppy.bot It's a direct replacement for ROS 2. We already have the OpenArm robot (https://openarm.dev) working on the platf…

hn · Vision

Show HN: FeyNoBg – Automatic background removal model and training library

Hey HN, I’m Shreyash from Feyn. We help companies build custom models from their data. Today, we’re releasing FeyNoBg, an automatic background removal model. Alongside it, we're open-sourcing NoBg, t…

hn · Vision

D-FINE-seg – detection, instance and semantic segmentation in one model

tc · Vision

Midjourney acquired the astrology app Co-Star

The AI lab Midjourney continues to expand its purview beyond image and video generation.

hf · Vision

Bringing Nunchaku 4-bit Diffusion Inference to Diffusers

hn · Vision

Show HN: Cactus Hybrid: We taught Gemma 4 to know when it's wrong

Hey HN, Henry & Roman here from Cactus. A small, on-device model is fast and private, but sometimes wrong, but frontier models are getting expensive pretty fast. So, we post-trained Gemma 4 E2B post-…

tc · Vision

AI and the rise of the universal entertainment app

Over the past decade, streaming platforms competed by dominating individual formats like music, video, podcasts, or audiobooks. Now, as AI makes it easier to create, organize, and recommend content, those distinctions a…

hn · Vision

Launch HN: Bloomy (YC S26) – AI-powered mastery learning for K-12

Hi HN, I’m Alex Southmayd, the founder of Bloomy ( https://bloomylearning.com ) – an AI-powered mastery-learning platform for K-12 students. Bloomy provides students with an AI tutor alongside adapti…

analytics · Vision

What is Seedance 2.5? ByteDance's 30-Second Single-Take AI Video Model

hf · Vision

Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers

hn · Vision

$100 AI Music Video: Claude Fable 5 vs. GPT-5.6 Sol

hn · Vision

Launch HN: Traceforce (YC S26) – Secure AI apps, one device at a time

Hey HN, we’re Xia and Varun, the founders of Traceforce ( https://www.traceforce.ai/ ). Traceforce provides visibility and control over AI apps such as ChatGPT, Claude etc directly on all devices (la…

tc · Vision

Reelful’s AI turns your camera roll into short-form videos for social media

The app is designed for people who want to create social content, but find traditional video editing tools too complex or time-consuming.

tc · Vision

Video-generation startup PixVerse raises $439M, valuation soars past $2B

With the cash, the company aims to expand its world model offering and reach customers across geographies.

hn · Vision

Show HN: YouTube Guitar Tab Parser

I created a simple CLI that turns a YouTube guitar-lesson video into a PDF of the guitar tab. There are services that transcribe music from Youtube videos into tabs, but they never work well enough f…

hn · Vision

Show HN: FableCut – A browser video editor AI agents can drive (zero deps)

tc · Vision

Google Photos adds a new AI ‘Video Remix’ tool

The feature can do things like apply cinematic relighting to brighten up a dark clip, swap out a plain background for something fun, or add artistic styles to videos.

tc · Vision

Why this CEO thinks video games make better training data than the internet

When it comes to achieving artificial general intelligence (AGI), large language models just don’t have what it takes. Models like ChatGPT and Claude are great at text, but they’re less skilled at understanding how thin…

tc · Vision

Midjourney wants Hollywood studios to reveal the details of their AI usage

As part of an ongoing legal dispute with three Hollywood studios, Midjourney is seeking to compel those studios to reveal how they use AI themselves.

hn · Vision

Claude-real-video - any LLM can watch a video

tc · Vision

Gemini’s personalized AI image generation is now free for US users

Google is expanding Gemini’s personalized AI image generation to eligible free users in the U.S., allowing the chatbot to create images based on your interests and data from connected Google apps.

tc · Vision

Apple Vision Pro exec is reportedly leaving for OpenAI

Paul Meade, the Apple vice president in charge of the Vision Pro headset, is reportedly leaving the company to join OpenAI’s hardware team.

hn · Vision

Show HN: Turn native language audio into flashcards and shadowing practice

Here is a tool I built initially for myself to help with my German and Greek language studies. It started as a hack for creating Anki cards from native language audio. It extracts the words, finds th…

hn · Vision

DiffusionBench: Towards Holistic Evaluation of Generative Diffusion Transformers

hn · Vision

Show HN: FastUbu – An Ultrafast Video Archive

Ubu is a 30 year old archive of strange films you'd usually only see in museums. I felt it was a perfect candidate for modern Midjourney-like performance. Really enjoyed using Cheng Lou's pretext and…

tc · Vision

Fika Jobs raises $4M to build a video-first hiring platform where AI agents interview candidates

Stockholm-based startup Fika Jobs is building a video-first hiring platform that combines AI interview agents with short-form video profiles, creating something that feels like a cross between LinkedIn and TikTok.

hn · Vision

AI is a mass psychotic delusion [video]

hn · Vision

Building a robotics research setup that lives next to my desk

Quick framing, since the post is long: I did robotic manipulation research at OpenAI from 2017–2020, and the tabletop setup back then cost roughly 10x this one and took a team to run. This project is…

tc · Vision

Snap spins off AI video team into new company, Dotmo, due to costs

The Snapchat maker is spinning off yet another internal unit. Dotmo will be comprised of current Snap staff who are leaving the social media company to focus on AI video development.

hn · Vision

I Think They [Anthropic] Are Lying to You [video]

tc · Vision

Cheaper, faster, and culturally aware, Avataar’s video AI is built for India’s scale

Avataar AI's distilled video model is priced at $0.005 for every second of generation

Browse briefing issues →

Get the briefing

The one story that matters, 5 headlines and the paper everyone's citing — every Tuesday, free.

Subscribe free