GPT-5

MediNAVI JAPAN

MediNavi JAPAN is a multilingual healthcare navigation platform for international visitors and students in Japan. The project was inspired by real challenges I encountered while working as a nurse at a clinic in Japan. Patients often struggle to identify the right medical specialty, find a nearby clinic that supports their language, or understand whether the facility is currently open and accepts walk-ins. MediNavi JAPAN helps users search for appropriate medical facilities by symptoms, specialt

AI Forecast Studio | Your AI Data Science Team in 30 Seconds

Built in 3 days using GPT-5.6 and Codex for OpenAI Build Week. AI Forecast Studio transforms business data into executive-level business intelligence by deploying an AI Data Science Team composed of specialized agents. Features: - AI Data Science Team (Atlas, Maya, Noah, Owen and Ava) - Statistical and machine learning model selection - Forecast Intelligence - Executive Decision Room - Team Meetings and collaborative reasoning - What-if business simulations - Executive reports and recommendati

SignalOS: AI Incident Replay | GPT-5.6 + Codex Demo

SignalOS transforms fragmented operational signals into a clear, evidence-backed incident replay and actionable response plan. This end-user walkthrough demonstrates: • Signal and incident monitoring • Evidence-backed timeline replay • Root-cause investigation • AI-generated summaries and recommended actions • How GPT-5.6 and OpenAI Codex supported the project Built for OpenAI Build Week. Try the live demo: #SignalOS #GPT56 #OpenAI #Codex #AI #IncidentResponse #Observability #Hackathon

OpenAI Agentic AI in Action Codex and the Future of Work

© 2026 Schmidt Special Competitive Studies Project, LLC. Licensed under CC BY-NC 4.0 (non-commercial use with attribution). To view a copy of this license, visit https://creativecommons.org/licenses/by-nc/4.0/

GPT-5's rollout fell flat for consumers, but the AI model is gaining where it matters most

OpenAI's GPT-5 has more than doubled coding and agent-building activity since its debut and driven an eightfold jump in reasoning workloads. Platforms including Cursor, Vercel, JetBrains, Factory, ...

Show HN: I rebuilt Microsoft Comic Chat's layout engine in one HTML file

I used Comic Chat in 1996 and never got over it. So when Microsoft put the source up last week, I wanted it back the way I remembered it: same engine, browser tab, no install, works on a phone.I do not write or read code. I'm the proverbial "product manager." So, I have to thank both GPT 5.6 and Fable 5 / Opus 4.8 (just before they released Opus 5 today). I aimed to stay faithful to the original as much as I could and extend it intelligibly.It's not IRC, but you can conn

Show HN: ScreenFocus – keyboard focus follows the pointer across Mac displays

I have a multi-monitor setup and spend a lot of time typing in Codex and other AI agents. When I move to another monitor to run a shortcut, focus often stays in Codex or another text app, so the shortcut runs in the wrong place.I asked ChatGPT if there was a macOS setting or free app for this. It couldn't find one that did exactly what I wanted, so I asked it to prototype one. It built a working version in one go.I then used Codex to add sensible defaults, focus protection, an optional disp

Ask HN: HotPin – lossless 120B MoE inference on 24GB RAM (CPU, 50 loc)

I'm a mechatronics designer with a background in control systems, robotics, PCB design, and embedded hardware. I design physical systems: motors, sensors, microcontrollers, and real-time control loops.I applied this design thinking to LLM memory management – and it worked.HotPin is a set of patches for llama.cpp that runs 30B–120B Mixture of Experts (MoE) models on far less RAM than their disk footprint, with bit-identical (lossless) output.Tested on an AMD Ryzen AI 9 HX 370 (Zen5, AVX512),

Ask HN: Which is the least sloppy and claudeism free model you have used?

I feel like recent models have been consistently getting more sloppy and increasingly claude-ism heavy (load-bearing seams galore) with every new release.I was hoping this trend would reverse in newer major version releases but I just tried out Opus 5 and it's similarly shit at writing. Fable is slightly better but not by far (and of course ridiculously more expensive). Adding skills helps but only mildly and not much at all for longer prose.Have you guys used any model lately that you like

Midjourney Releasing v8.2

The V8.2 making it the default model on Midjourney. This is a release focused on aesthetics, personalization, and image quality. Our new style is more creative, bold, edgy and fresh and personalization works better than ever.

Show HN: Mumble Dictation – local dictation that learns your vocabulary

Hey HN! It’s Wen from Narya.ai. We build Mumble Dictation, a local-first and privacy-focused dictation app on Mac.Homepage and Download: https://heymumble.com/dictationYouTube:https://youtu.be/Bma59TrR62c?si=DM-LyjZY5GZ9-f64Two things kept bugging us about the dictation products on the market:1. Most of them ship your audio to the cloud, and those who is claiming to be private are often just a thin wrapper of cloud APIs with a privacy policy.2. Those general-purpose

Show HN: CobaltCode – Dedicated persistent computer for Codex

Hi HN,I'm building CobaltCode, a platform for running coding agents inside persistent development environments.Basically I was fed up with git worktrees and having to run everything locally, usually one at a time so built CobaltCode.ai. Each task gets an separate vm with its own repository checkout, dependencies, running services, agent session, previews, and persistent state. It can also run your app inside the vm so you can preview different features in parallel.You can leave a task, resu

Show HN: Agent in 9 Lines Python

I asked myself: what would a minimal implementation of an agent look like?Something that works out of the box, is a real agent with tool calling, but without 1000s of lines of code, without dozens or hundreds of npm or pypi dependencies. Something with just a few 'essential' features (not a whole kitchen sink that most agent harnesses come with nowadays).An implementation close to pseudocode that you can look at in one page, everything there at a glance, no scrolling.This is the agent.

Show HN: Millwright – Rust-based, self-hosted LLM router

Hey HN,With the news of OpenRouter possibly being acquired and proliferation of hosted LLM routers (i.e. Ramp Router, Vercel’s AI Gateway), I saw the need for a self hosted solution focused on cost savings, transparency, and performance. So, I built an open sourced router with a simple CLI interface that can easily sit between coding agents and GenAI workloads.For the curious and lazy, at the moment, Millwright has the tools for,- Providers: OpenAI-compatible APIs, Anthropic, Amazon Bedrock- Rou

Show HN: Use your flight-sim gear as a Codex Micro

When I saw the OpenAI Codex video for the (cool-looking) Codex Micro, I looked down at my Virpil CM3 throttle collecting dust.So I just asked Codex desktop to connect it for me-- and it turns out the latest desktop app has the hotkeys/bindings defined for the Micro already. In a couple of hours I had the core options bound and over the weekend I got to full LED status control.The code is Windows stuff and probably not the hardware you have on hand, but that almost doesn't matter anymor

Ask HN: What do you do when your AI agents are working?

Maybe a dumb question but I don&#x27;t think I&#x27;ve ever experienced more down&#x2F;waiting time in my career so I&#x27;m much more prone to opening YouTube, HN, etc. despite not really wanting to.<p>Most likely the answer is to just sit and do nothing. But have you discovered an approach to making agentic coding, etc work for you without feeding into distraction and attention deficit?

Codex Is Down

Being investigated as we speak. https:&#x2F;&#x2F;status.openai.com&#x2F;

Companies are optimizing models for specific benchmarks

Openai is optimizing for gpqa diamond and anthropic is optimizing for humanity last exam. gpt 5.6 wins on gpqa and opus 5 wins on humanity last exam

Judge approves Anthropic's $1.5 billion settlement with authors

A federal judge in San Francisco has granted final approval to a $1.5 billion settlement between Anthropic and a class of ...

Anthropic launches Claude Opus 5 with efficiency, safety improvements

Anthropic PBC today rolled out a large language model called Claude Opus 5 to its chatbot service and developer platform. The ...