Halo
Post-training framework for open-source models. Up to 2.8x TRL throughput, less peak memory, weights stay HuggingFace-native.
Feature a build or sponsor the directory. DM @pranavmore69
Models, wrappers and small products that do something a frontier model could not do last year. We only list ones you can try, not launch-thread vapor. 211 builds on this page, each with the original post.
Post-training framework for open-source models. Up to 2.8x TRL throughput, less peak memory, weights stay HuggingFace-native.
Open 151M decision model on ModernBERT. JevBench #2 on public tasks after the inference-engine fix. Weights and a WebGPU playground.
Asked Muse to research itself, Amazon vs Shopify, and which stocks win if agents start shopping. It built the page.
Open Apache 2.0 System One model. ModernBERT, typed choice/score/noul in one pass. Fine-tune notebook on Kaggle T4s. Not Jev.
Jared Palmer's open Jev-like family on Qwen3.5. 0.8B, 4B, 9B. TypeSafe SDK with a base_url change. Train on one H100.
Open decision model that reads option logits from local Qwen-class weights. Browser demo at openjev.com. TypeSafe SDK compatible.
Laya on Apple MLX. Under 1GB RAM, 7 to 14ms on M3 Max. Snake at 60 decisions a second, no PyTorch, no cloud.
4B, 9B, and 35B-A3B decision family for browser agents. ~26ms median. TypeSafe SDK. Weights on Hugging Face.
Open System One model in the same class as Jev, Kev, and Laya. State plus typed questions in, calibrated probabilities out.
scikit-learn workflows on Apple Metal GPU. Change the import, keep fit/predict. Falls back to CPU when it has to.
Meta's personal AI agent. Own Secure VM, own browser, does the work across apps. Powered by Muse Spark 1.3.
Meta Superintelligence Labs flagship. Agentic and coding model in Muse Code and the Meta Model API.
A system one model like Jev. Typed answers with probabilities, no generated text. pip install seacat or npm install seacat-ai.
Realtime support triage with Jev routing tickets, then Cerebras drafts the reply. Typed decisions, no text to parse.
On-device LLM image understanding in Unity. Offline OCR by describing the target in natural language.
Poor man's Jev on local models via omlx. Same API shape as Jev, for people without access yet. From GitHub Next.
Chrome extension that listens to YouTube audio, detects sponsor segments, and skips them in real time. Open source, BYOK.
Local-first AI data analyst. Your data stays on your machine. Only column stats leave.
TypeSafe's System One model. Typed decisions, no chat. Trained with RLCD. Output tokens free.
Playable Jev demo. A needle that panics in realtime. Possibilities, not text.
Browser horror game: they only move when you cannot see them. Desktop and headphones. GPT-6 Astra plus fal.
Trading bot with Jev. Buy or sell from a price feed, then places orders on Monad via Kuru every 300ms.
302-neuron C. elegans wiring trades simulated SOL. Silence a neuron and replay the market.
Open-source GARCH trading framework: Pine Script indicator, agent skills, and a Claude plugin.
Local LLM gate that hides sensitive info before it reaches online model providers.
Bring Jev-style decision scoring to existing LLMs. Calibrate on your data, set a review threshold, no fine-tuning.
Gaussian splat browser game in a World Labs Marble scene. three.js, nav mesh, light probes. First-party, with a walkthrough.
Two Astra prompts, a reference video, and a product page. Playable React fold demo.
Sunset to nightfall to moonrise to Milky Way, all JS, no images. First-party Astra build.
One video of a human hand becomes a real2sim pipeline. Tracking, IK to a 44-DoF hand, grasp refine. First-party, OSS.
Open-source FPV flight controller. ESP32-S3, EasyEDA project and gerbers. Designed with Codex and GPT-6 Astra.
World action model that lets robot policy scale with in-the-wild human video. CoRL 2026. Code and data are public.
NVIDIA inference stack for MiniMax-H3. Five seconds of video in 1.65s on 8x B300. Faster than playback. Apache 2.0.
Free agent skills that force a straight verdict, plain steps, and a next action. Built by a non-engineer running a software company.
Open-source ComfyUI nodes and workflows for Reactor endpoints, including MiniMax FastH3.
HeyGen open-sourced real-time AI avatar demos with OpenAI GPT-Live-1, LiveAvatar, and HyperFrames.
MXFP4 vLLM build for Radeon AI PRO R9700. NVFP4 Qwen3.8 27B at 5809 tok/s prefill.
Buy-once social scheduler. Unlimited channels and posts. Talk to it from Claude or ChatGPT.
GPT Astra / Claude skill that locks one face, voice, and light across many videos.
Generative video classroom on fal H3 Max. Ask a question, watch an animated explainer, queue follow-ups. OSS, add your own keys.
Turns a favorite song into a custom blind-box toy. Built with Reve. Playable.
The open-source X ranking code. Daily updates, and the first public PR is now live on X.
A live Pangram AI-detection badge you can drop on a site. First-party, with a generator and the repo.
Agents turn physical experience into executable world code. Models, paper, and a project page from MirroS.
Anthropic's public Mythos-class model. Long-horizon coding and research, cheaper cache reads. Live everywhere today.
Same weights as Fable 5.1, fewer safeguards. Restricted to cyberdefenders and life scientists via trusted access.
Fable 5.1 on Google Cloud Agent Platform. First-party from Google Cloud.
Claude Security scans now run on Mythos 5. Public beta for Claude Enterprise. First-party.
Mythos Preview found new attacks on HAWK and a reduced AES. First-party research from Anthropic.
A from-scratch PyTorch reconstruction of what Claude Mythos might look like. Looped transformer, MoE, MIT.
Fable 5.1 decoded Sir Thomas Urquhart's 1653 cipher, unsolved for 373 years. First-party from Vals AI.
Zero-config Spring Boot starter that adds AGENTS.md support to Spring AI apps. First-party.
Local-first annotation workspace. Non-destructive image pipelines, synced views, and model feature inspection.
Z.ai's natively multimodal Flash model. 320B total, 18B active, 1M context, MIT weights. Previewed as Ox Alpha.
The full GLM-5.3 coding and cyber-defense model is now downloadable on Hugging Face. Same base as 5.2, post-trained.
Alibaba's multimodal MoE and a preview of Qwen4. 125B plus 51B n-gram embeddings, 6B active, 262K native context. Open-weight.
Unsloth 3-bit GGUF so GLM-5.3-Flash runs on 128GB RAM. First-party local pack.
Unsloth 2-bit GGUF of the full GLM-5.3. 1.51TB down to 239GB, for a 256GB Mac or mixed RAM/VRAM.
Official MIT harness from DeepSeek. Everything is a plugin: models, tools, sandboxes, loops, and UI. Run with npx @deepseek-ai/dsh web.
DeepMind's harness for autonomous research runs that last days to weeks. Used internally on Simply, now public.
Inference runtime, local API, agent loop, and training provenance for Dot's 14B model. Company open-sourced the engineering layer.
High-angle robust fast face alignment. Whole-head dataset and training pipeline, first-party from PINTO.
Teknium's stand-alone Hermes Agent plugin. Search the web as it was on a given date, like a Wayback Machine for agents.
AI-native CAD in the browser. The author designed a physical coaster in it, then manufactured it.
Simon Willison's page that paints LLM tells. 38 patterns now.
ROBOTIS OMY arm: Gemini reasoning, GraspNet 6-DoF poses, collision-free pick-and-place on ROS 2.
All 63 US national parks scored on remoteness, scenery, elevation, biodiversity, plus camping and directions.
Desktop app that runs Claude Code, Codex, Cursor, and 26 other coding-agent harnesses in one place.
Sandboxed AI harness in Zig. Landlock, eBPF, and namespaces isolate tool calls. WASM plugins over IPC.
Turn any number of photos into one 3D surface with Flow Matching. Code and model are public.
Internal agent apps Rome runs the company on, open-sourced so you can see how they actually work.
Installer that puts GLM-5.3 next to Claude models inside Claude Code, via OpenRouter.
Turn a repo into a knowledge graph and hand it to Codex, Cursor, or Claude Code over MCP. Runs in the browser.
Blur faces and objects on a Mac without uploading the video. YOLO on the Neural Engine, tracking so a missed frame does not leak a face.
Generative art where the model writes JavaScript in an isolated worker instead of getting a tool per shape.
Speaks numbers and processes, draws charts and diagrams while you talk. Local Whisper. Cannot invent a number you did not say.
Decoupled humanoid whole-body controller. Train on one GPU, drive it in the browser, zero-shot onto the robot.
NVIDIA SRL's robot control stack, rebuilt so the hardware is actually easy to set up. pi0.5 doing the hello-world task.
Tag an AI in Slack the way Claude Tag works, but pick any model. Live with 10 teams.
Pixel-space Diffusion Transformer for dense 3D point maps. Code, weights, and training pipeline are public.
Spy-simulator in the browser on real data: planes, ships, satellites, traffic cameras. Talk to the planet. MIT.
vLLM plugin that fixes compaction and thinking-mode bugs when running Qwen 3.8 27B as a local Hermes agent.
RL harness that treats ASIC design as a hill-climb. RTL and GDS are public. Write-up at luoluo.ai.
iOS app to download open models from Hugging Face and chat with them on-device. No API keys, no terminal, no subscription.
Thin harness so an agent can drive a full Omarchy desktop: Hyprland, AT-SPI, real browser, files.
Self-hosted knowledge vault for Hermes: ingest articles, YouTube, RSS, bookmarks, then search them from the agent.
Open-source AI IDE for Claude Code, Codex, Grok, Cursor, OpenCode. Parallel agents, each in its own worktree.
Tiny macOS menu bar that shows how much Codex you have burned without digging through menus.
Harness that gives every AI bot its own directory, browser profile, cookie jar, and environment. Grants expire; the door closes.
Dashboard for multi DGX Spark setups: inference throughput, GPU processes, and more.
Turns each agent rule into a function that checks the agent's work, instead of hoping it follows the prompt.
Webcam emulator inside a spreadsheet, built in Google AI Studio on Gemini 3.7 Flash. Export images and CSVs.
Open-sourced detector for AI-slop websites. Add a pattern, improve a detector, or build on top of it.
Inference stack for Qwen3.8-27B on a single RTX 3090: 195k context, 82 tps single request, 672 tps peak.
Day-one Qwen3.8-27B deploy: NVFP4, 262K context, one DGX Spark, one command. Full reproducible stack.
27B on two consumer 3090s at ~75-85 tok/s with NVLink. Tool calling, reasoning, vision, MTP. Full deploy open-sourced.
Poole Adaptive Reasoning Stack as a skill: generate exact binaries, research, and audit with PARS.
Open-sourced agent memory. Mid-session compression cut 61% of token usage in their case.
Open agent memory on Cloudflare Workers and Durable Objects. Profiles, extract, recall. One-click deploy. MIT.
OpenRouter for tools: 2,600 agent-friendly APIs, pay per call, 0% markup, search by task. Open source.
First-person roguelike deckbuilder: defeat monsters, harvest parts, cook meals. Playable demo after two months of building.
85-line finger-tracking toy for Physical AI video cleanup. Claude Code plus MediaPipe.
Open-world Bucharest satire from an AI Game Jam. Playable plus repo.
Resume classifier. Streamlit app plus GitHub.
Type an order like a text, get an AI invoice, payment link and receipt.
YC-backed open-source computer-use alternative to Claude Cowork.
Harness that automates 'keep going' on open math problems.
Open-source motion-graphics framework and skills from HeyGen.
Cross-institutional patient memory layer on Cognee. Hackathon.
Universal model router for AI coding agents. 11k GitHub stars.
Persistent local memory for coding agents. One Python file, no dependencies, explainable recall.
ChatGPT clone after six GenAI cohort classes.
10 apps/tools for AI-native filmmaking.
Full ESP32 AI project, open-sourced.
Open-source AI repo from the same dump.
AI UI elements for Svelte. Live demo.
Can a VLM act through a body? Open-source benchmark for action intelligence, decoupled from motor control.
25 small biology apps for day-to-day experiment planning, built in a couple of hours on GLM-5.3.
Reason around a question without context-window limits, and visualize the knowledge graph.
Open-sourced genetic / idea-virus project.
AI coding tool with Gemini 3.7 as the default. Bring your own Google AI Studio key.
Public engine to stream DeepSeek V4 Flash experts on an 11GB laptop. LRU under one layer is 0% hits, so they published it.
PyTorch from scratch in C, to understand how it actually works.
Minimal chatbot template: new components, typeset, tool calling, HITL questionnaire, message parts. One-click deploy to Vercel AI Gateway.
oh-my-pi extension for best-of-N + self-verify: sample several solutions, rank them with the same model as verifier.
Shared local brain for Claude Code, Cursor, Codex and OpenCode. MCP plus CLI so agents stop redoing each other's work.
On-device LLMs in Flutter on Snapdragon. Gemma and Qwen via Qualcomm QNN, Hexagon NPU, local streaming.
Personal search over X, YouTube, RSS, Bluesky and Mastodon. Web UI plus MCP so agents can query it.
iOS app that scans your face, scores your skin and gives a shareable routine.
A single-file HTML canvas Breakout game written end to end by a local Qwen 3.8 27B through Hermes.
AI security platform built for the new era of AI threats, defending companies against AI-powered cybercrime. YC S26.
AI search, product recommendations and customer support for online stores. Ranked second for AI Search on the Shopify App Store.
A 24/7 agent for solo game devs. Watches Steam, itch.io and Discord, replies to players, files bug tickets and drafts weekly patch notes.
Research assistant for investors, built to dig through company filings and documents for deeper insight.
An AI-native Layer 1 experiment on the Canopy testnet, giving autonomous agents their own infrastructure and economy.
Open-source AI agent that reviews your code and auto-fixes it.
MIT tmux pane that renders what your coding agent can see and you can't: images, video, PDFs, diffs and X cards. Bash 3.2, nothing leaves the machine.
MCP bridge so ChatGPT can operate your real Mac: shell, files, PTYs, background jobs and read-only Codex history. MIT. No model calls.
AI app for SMBs to chat with customer feedback, conversations and sales. Launched on Product Hunt.
Turns any question or block of text into a video explainer instantly. Also ships as a Chrome extension and as Claude and ChatGPT plugins.
A custom token launchpad paired with a Claude pipeline that turns study notes into textbook-style blog posts.
A 15-second fictional luxury fragrance campaign generated entirely with Grok Imagine from two reference images.
AI agents that crawl the App Store and Play Store around the clock and deliver MASVS and GDPR digests the moment a new app ships.
A trading bot built with Claude Code in two days that scans 50+ markets for arbitrage. The post claims it turned $68 into $750k.
A $2,300 local AI rig: two RTX 3090s for 48GB of VRAM, a dual-Xeon board and 28 cores, running 70B quantized models fully offline.
Closes the gap between the idea in your head and something people can actually use and ship.
Continuously tracks the biggest AI stories, funding rounds, partnerships, model releases and infrastructure news.
Open-source AI agents that run a full pentest end to end, from recon through exploit, reverse shell and pivot, then write a client-ready PDF report.
An AI cowork harness tested heavily against open-weight models. Runs well on DeepSeek V4 Flash and Meta’s Muse Glimmer 30B, which is agentic out of the box.
Two cancer tools from a practicing oncologist: an AI tumor board for clinicians, and a multi-agent patient-facing tool that answers the questions doctors have no time for.
Fires a forked thread and returns an improved prompt, with a diff view component to show what changed.
Human heart anatomy site built with Gemini 3.7 Flash. Explore the heart in 3D and learn what each part does.
An agent built with Claude, GitHub and Notion that finds recently merged PRs, summarizes what shipped, and posts a fresh digest to Notion every Monday.
Splits any AI-generated image into editable layers for subject, text, logo and background, so each piece can be restyled independently.
Generates software installation instructions for application owners and packaging teams, cutting the manual documentation work.
Turns a firm's design workflows into software. Describe the workflow, define the rules, and get an app that runs on your building model.
A 3D Japanese tofu delivery drift game built entirely on Crayon Pro, an AI game engine.
Z.ai’s new coding and cyber-defense model, post-trained on a 743B base. Agentic coding jumps over GLM-5.2 and it is live now via GLM Coding Plan and ZCode.
Free, open-source Granola alternative that actually looks good. Bring your own OpenAI and AssemblyAI keys, optional R2 storage.
Free, offline, open-source workpaper binder for accountants. Agents connect over local MCP, every change is journaled, and humans still sign off.
Basement-built autonomous rover: custom BMS PCB, STM32, Raspberry Pi, and a neural net trained to push objects into a goal zone.
Agent tools for any SDK that live-refresh and can be searched or described, plus a durable agent mailer along the lines of Oban for JS.
Ten fresh indie launches in one roundup: Procright, Anomaly Ai, Bloom, Popform, Max, Bon Split, Archmaster, Flexify Suite, Kage and Roman Names.
Free open-source Postgres intelligence for agents and apps. Finds unused indexes, slow queries and drift. Written in Go. No dashboards.
Fully shipped AI agent with persistent memory built in less than a day.
100% open-source meeting assistant built with Rust + Tauri. Transcribes 4x faster with Whisper/Parakeet and runs completely locally.
Turns any PDF, EPUB, DOCX or Markdown book into a ready-to-use Claude Skill that loads only the relevant chapters when needed.
AI game creation platform that turns text prompts into playable browser games.
Native inference engine optimized for DeepSeek V4 and GLM models with support for Metal, CUDA and ROCm.
Local-first visual LLM workspace where graph edges control model context, with branching, merging and document extraction.
Native iOS AI companion with animated VRM character, voice/text chat, local history and camera context support.
High-fidelity image-to-3D generative model using compact structured latents.
Multilingual speech recognition, text-to-speech and zero-shot voice cloning WebUI.
Open-source CLI agent for authorized pentests. Plan, act, observe, verify, report. Local or hosted LLMs. Burp + MCP.
Grokathon project that lets people speak without using their mouth by combining brain-sensing tech with Grok Voice.
Deep dive into the lesser-known but extremely powerful features SpaceXAI has shipped in Grok Build.
Human-in-the-loop teleoperation infrastructure aimed at creating an open, distributed learning network for robots.
Persistent Grok-based AI agent with its own wallet, identity and ongoing lore around the “Year of the Singularity”.
Upload your photo and measurements to see exactly how clothes will look and fit on your body.
Open-source repo that grew from an 11-agent experiment into 18 reasoning personas with multi-provider routing and confidence-weighted verdicts.
Program built with Claude Code that interviews job applicants over chat, scores fluency and thinking, and flags likely AI-assisted answers. Live in production.
Fully open-source agent operating system that runs on mobile, desktop and web with voice + chat control.
Vibe-coded tool that redesigns your room using real furniture you can actually buy, with a full shopping list.
AI that has its own phone number and email. Makes calls, waits on hold, and books follow-ups. #1 on Product Hunt.
Coding agent you speak to instead of typing. Say what you want and it ships in the background.
Agent that schedules meetings and books reservations 100% over text message. No app required.
Fully offline portable translator running Gemma on a Raspberry Pi 5 inside a custom 3D-printed case.
AI tool focused on meeting intelligence and action extraction.
Voice agent platform for building conversational AI experiences.
Command-line AI agent that works directly with Gmail.
176KB C99 engine that runs the 2.78 trillion parameter Kimi K3 on a single CPU with 8GB RAM by streaming experts from NVMe. No GPU, no framework.
Open-source self-improving coding harness where sub-agents are function calls in a persistent kernel. Scored 95.5% on ARC-AGI-3, above the human expert baseline.
Kenton Varda's remake of his startup Sandstorm as an AI agent workspace with connectors, built entirely on Cloudflare Workers. #1 on Product Hunt.
Voice AI product that gives live coaching reports (talk ratio, objection handling, buyer verdict). Reached Top 5 Product of the Day.
Rust open-source parser that converts PDF, DOCX, PPTX and 10+ formats to clean Markdown in under 5ms, built for agents. 12.5k stars in days.
Secure enterprise-grade platform that lets companies vibe-code production software inside their own AWS environment.
AI coding agent that takes a single prompt and ships a complete project.
Fully open-source framework that turns an $8 ESP32 into a real-time autonomous voice assistant with local wake-word detection.
Open-source control plane for AI agents. Defense by default. Used by thousands of devs a month.
AI product frequently highlighted in maker shoutouts for its capabilities.
Tool highlighted by makers for its usefulness in the current AI building wave.
Tiny open-source AI model running fully offline on a few-dollar ESP32-S3 chip at ~9.5 tokens/s. No cloud needed.
Open-source tool that gives frontier models effective 100M–1B token context by using a graph-based retrieval system.
Full Physical AI data collection and processing platform for training robots from simulation and egocentric video.
Custom local AI running on an $8 ESP-32 that tracks humans with zero computational delay.
Open-source tool that turns plain English descriptions into production-ready CAD files (STEP, STL, URDF, etc.) for robotics and mechanical design.
AI-assisted slot game made by a complete beginner that reached #2 on Steam demos with 80%+ positive rating.
Open-source desktop AI coworker that delivers finished work: polished documents, Slack messages and calendar updates, with 35 connectors and permission tiers.
Game engine built specifically for AI agents. Small enough for agents to understand the whole codebase and generate complete playable games.
Platform that lets you build apps and agents with a single prompt.
Autonomous AI agents running on the Arc network.
Introducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of stock TRL with less peak memory, while models stay in their native HuggingFace format. Star us on GitHub: github.com/whitecircle/ha…
Did my openJev Verdict(Oss jev made on just 151M param base model) hit #2(possibly #1 even?) on JevBenchmarks> When JevBench v1.2 evaluated openJev Verdict earlier this week, we were sitting at #11 with a score of 66.2. I audited the benchmark run this morning and realized the Show more
Did my OSS-Jev called Verdict, beat Official Typesafeai 's Jev ?, and Laya (the best performing OSS Jev) on their own benchmark? When I first tested Verdict 1.0(my earlier attempt of oss-jev) earlier this week, it scored 26.10% accuracy. Uniform random guessing on that split is
I asked Muse to research itself — its top use cases, the Amazon vs. Shopify fight, and which stocks win or lose as AI agents take over shopping and travel. Then it built me a whole page about it. 🔗 muse.ai/s/muse-and-age…
Asking everyone to fine tune github.com/NandhaKishorM/… Just use any kinda agentic harness like claude code, codex etc. Change the code and use free kaggle GPUs to train your own use cases. Base model has limitations
Kev has now been refactored on top of Qwen3.5. New checkpoints are now available at 0.8B, 4B, and 9B along with a fine-tuning script you can use with @modal. github.com/jaredpalmer/kev
UPDATE: Kev-0.6B, 4B, and 8B are now available. Kev is a family of small open source Jev-like decision models you can train and run yourself. This new family is based on Qwen3 using the same LoRA + small pointer head technique as before, but scaled up. Out of domain, on data
SemIf update today: merged Apple Silicon/MPS support (GitHub: dp-IED), a 27B EXL3 bridge hitting 95.8% balanced accuracy (@jkyamog) , and proper temperature calibration (GitHub: samarthpatel24) 📊 Check it out at openjev.com, GitHub.com/theoleecj/semif
介绍比Jev快50倍,在你设备上跑的laya-mlx! 只在你的设备上占用最高1G内存 Laya是一个开源的类似于Jev的,基于文本输出概率的分类系统 我将其移植到MLX,并且做了一些性能优化! 视频中就是这个模型在我的本地M3Max上玩贪吃蛇 这个模型能够以每秒决策60次的速度玩贪吃蛇! github.com/mizorewww/laya…
🚀 APUS-OpenJev: I built a 4B/9B/35B-A3B model family for browser agents and business workflows. ⚡ ~26 ms median response 🎯 82.2% vs Jev’s 77.0% 🧠 Selectable compute depth 🔌 TypeSafe SDK integration Models (huggingface.co/collections/ap…)
I just open-sourced MetalML. MetalML is a open-source library that brings up to a 15x speedup on native Apple Metal GPU acceleration to scikit-learn style machine learning workflows in Data Science. MetalML dispatches eligible work through native Metal kernels + Metal Show more
1/ today we’re releasing muse spark 1.3—available in muse code & the meta model api. this is our most capable model yet—frontier performance almost too cheap to meter. much stronger at agentic and coding with better usability. we think users will really notice the jump.
Today I shipped SeaCat. It’s a system one model like Jev. Libraries available: - pip install seacat - npm install seacat-ai Try it out: seacat.dev
i just shipped a small repo that will help you understand it perfectly.its a Real-time support triage + response bot using Jev and cerebras github.com/TheEleventhAva… PS if you like it,approve my entry to the hackerhouse. I applied with rewantgoenka87076@gmail.com
Happy to finally share this! I've open-sourced the Unity sample shown in the demo video I posted yesterday. It runs an LLM fully on-device, doing image understanding entirely inside the phone. No network, no cloud. [GitHub] github.com/TakashiYoshina…
Today’s prototype: fully offline, LLM-powered OCR running entirely on a smartphone. Unlike conventional OCR, there’s no need to position text inside a fixed ROI. Simply describe the target in natural language, such as “the sign on the left,” and the app extracts its text. #gemma
github.com/githubnext/loc… Not everyone on the team has access to Jev yet. Spent a morning cobbling together a poor man's Jev on top of omlx for local use. Benchmarked and eval'ed a variety of models including diffusiongemma and a variety of autoregressive models (Qwen MoE, Gemma 4 Show more
Just trying out Jev, I made a Chrome extension that: - Listens to your YouTube audio (optional) - Detects if it gets to a sponsor segment - Skips it ➡️➡️➡️ - All in real-time while costing ~$0.005 per video Prototype project, BYOK, open-source: github.com/trungdq88/yout…
Hello everyone, Today wektor is finally launched As a student solopreneur, i would love genuine feedback from u all . U can try it here: wektor.vercel.app Local-first analysis.Your data never leaves your machine. Only column stats do. #data #privacy
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x Show more
Spent this morning with Typesafe’s Jev - a model that returns possibilities, not text. I made a needle that panics in realtime ask-jev.vercel.app
it's insane what you can now do with GPT-6-Astra + fal i made this browser horror game inspired by the weeping angels in Dr. Who: they only move when you can't see them Try it. Desktop + headphones 👁️ weeping-angels.vercel.app
I built a trading bot with Jev! Jev decides if it should "buy" or "sell", given the price feed of an asset pair, and executes real trades. It uses Monad to place the orders on Kuru's on-chain order book in every 300ms block. Demo link → jev-trader.vercel.app
I built NEMA. 🪱 302 neurons. One market. Published worm wiring powers a neural model that trades simulated SOL. Watch its activity, silence a neuron, and replay the market to see what changes. An 8-bit worm with an open-source nervous system. github.com/ashleyotooliga…
I open-sourced a real quant trading framework to help you make money. The GARCH Method is used by real quant desks every single day. GitHub contains: Pine Script indicator for TradingView, agent skills, Claude Plugin (+ instructions). 100% free 👉 github.com/milesdeutscher…
yup, I built something similar to use local LLM to hide sensitive info from online LLM providers github.com/madeye/sealgate
Jev made me ask: can we bring the decision-model approach to existing LLMs? So I built DecisionBridge. Your choices → model scores → your code decides. Calibrate on your data. Set a review threshold. No fine-tuning. Code: github.com/grishahq/decis…
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x
I made a little gaussian splat game demo set in a world created in World Labs Marble, using three.js + spark + crashcat + navcat, and I made a video about it! On the @theworldlabs youtube channel now: youtu.be/TiLRhSdGsqY + project github: github.com/isaac-mason/sc… My Show more
astra. 2 prompts. gave it a reference video and the product page. insane lmao react component btw github.com/jal-co/iphone-… iphone-duo.vercel.app
i had to really push astra on this, but it did end up coding this nice transition from sunset -> nightfall -> moonrise -> milky way there's not a single image in here. it's all just JS code. i'm mildly impressed repo: github.com/kunchenguid/fl…
Video in → physics out. 🤖 I gave GPT-6 Astra one video of a human hand. It wrote this entire real2sim pipeline itself — hand tracking, IK retargeting to a 44-DOF dexterous hand, grasp refinement. Zero code from me. Open-sourcing everything 👇 github.com/Hu-xiao-max/de…
open sourced the FPV flight controller hardware ESP32-S3, EasyEDA project + gerbers in Releases. designed with Codex + GPT-6 Astra. PX4 bring-up still WIP, board already ordered. github.com/hx23840/S3-PX4…
Reworked the FPV PCB. Gave GPT-6 Astra a few clean reference boards, then redid IC placement and routing. V1 → V2 below. Ordering the board now.
EgoWAM has been accepted to #CoRL2026 @corl_conf 🥳🥳 See you all in Austin!!!! We’ve also open-sourced all the data and code: 📊 Data: partners.mecka.ai/egoverse?task=… 💻 Code: github.com/GaTech-RL2/Ego… If you run into any problems, please submit an issue on GitHub. We’ll do our best to Show more
Egocentric human data is abundant, but human motion is not always positive supervision for robot policy due to embodiment gaps. Naive BC co-training can HURT performance ☹️. 🌟Our key finding in **EgoWAM**: the state-prediction branch of a World Action Model effectively bridges
🚀 Sol-H3: @MiniMax_AI H3 Video Generation Faster Than Playback 🤩 Five seconds of world. 1.653 seconds to infer. We’re releasing Sol-H3, our fastest end-to-end MiniMax-H3 inference stack yet. On one 8× NVIDIA B300 Blackwell system, it generates five seconds of 1344×768 video Show more
I am not an engineer. I run a software company and use coding agents daily. Their output is written for engineers, so I built Open Steps. Free skills that make an agent give a straight verdict, plain steps and a real next action. 250+ stars in a month. github.com/kharmanskyi/op…
Open-sourced a bunch of @ComfyUI nodes and workflows for @reactorworld endpoints. This includes their endpoint for MiniMax FastH3. The repo still needs cleanup but it's ready to use. Enjoy: github.com/stefanionescu/… Might be useful for others participating at Worlds Hackathon Show more
Real-time AI experiences are now solved AND open sourced. We worked closely with @OpenAI to build the best framework for it, combining GPT-Live-1, @TryLiveAvatar, and @HyperFrames_ Check out the demos below and show us what you build with it. github.com/heygen-com/liv…
NVFP4 Qwen3.8 27b - AMD R9700's running @ 5,809 tok/s Prefill and 276 tok/s decode. This leverages the MXFP4 fast path I built on vLLM Radiance. github.com/GGZ14/vllm-mxf… #AMDAI @AIatAMD #localLLM #amd
introducing: SocialQuack.com stop paying monthly for your social media scheduler unlimited channels, unlimited posts, unlimited joy buy once, own forever connect to Claude or ChatGPT and just ask it: "brainstorm & schedule a month of posts about my startup [your Show more
Good prompts are rare and making them sucks so i shipped a GPT Astra / Claude skill Can make you 50 of these in under 30 min ads? lock a character or product same face. same voice. same light. across 50 videos free: one-face-lock.vercel.app
So delicious 😋 Seedance 2.5 Prompt: Create a 30-second fast-paced cinematic Japanese anime cooking video showing the preparation of authentic katsudon, entirely from the text description below. IMPORTANT: Do not display, recreate, trace, reference, or imitate any storyboard,
Open sourced it here, just add your own @fal + LLM API key to generate videos locally. Also updated to use the new H3 Max Turbo mode that just launched, which is 2x faster and half the cost github.com/internetphysic…
I built a generative video classroom with @fal H3 Max Ask about any concept, and in just a few seconds it'll begin playing an animated explainer video taught by Tung Tung Tung Sahur. You can also queue up new videos and ask follow-up questions while you're watching existing
I love the design of 3 things very dearly: spray paint cans, album art, & collectibles ..so I built a little site with @reve that turns your fav song into a custom blind box toy sound on for more fun :) rattlers.vercel.app
Today marks 2+ weeks of daily updates to the open-source X algorithm. We’ve also integrated a first public contribution — a small update based on a Pull Request submitted to our GitHub repository, now live on X: github.com/xai-org/x-algo… Thank you to everyone submitting PRs.
made a @pangram badge thingy because i couldn't find one site: pangram-badge.ryanaque.com repo: github.com/schmayterling/… blog about it: ryanaque.com/blog/making-a-… badge below: pangram.com/history/0b1997…
Big thanks to @_akhaliq for sharing Code-as-World 🌍 We’ve open-sourced Code-as-World — a step toward turning physical experience into structured, executable knowledge, and ultimately toward Physical RSI. 🚀 Code & Models: github.com/mirros-lab/cod… Blog: mirros.ai/blog/represent… Show more
Code as Worlds Agentic Discovery of Executable World Representations for Physical Reasoning paper: huggingface.co/papers/2608.27…
We’re introducing Claude Fable 5.1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge work.
Claude Fable 5.1 is available everywhere today. Claude Mythos 5.1, our model for cyberdefenders and life scientists, is available through trusted access programs. Read more: anthropic.com/claude-fable-a…
Claude Fable 5.1 is live on Agent Platform 🚀 Put frontier intelligence into production across coding, scientific research, and enterprise workflows on Google Cloud. Try it today → console.cloud.google.com/agent-platform…
Claude Security scans now run on Claude Mythos 5, available today in public beta for all Claude Enterprise customers. Put our most capable security model to work on your codebase, no separate model access needed.
New Anthropic research: Discovering cryptographic weaknesses with Claude. Claude Mythos Preview has helped our researchers find weaknesses in cryptographic algorithms—the mathematical methods that are used to keep data private. Read more: anthropic.com/research/disco…
Introducing OpenMythos An open-source, first-principles theoretical reconstruction of Claude Mythos, implemented in PyTorch. The architecture instantiates a looped transformer with a Mixture-of-Experts (MoE) routing mechanism, enabling iterative depth via weight sharing and Show more
Using Fable 5.1, we’ve elicited a decoding of a cipher that has been unsolved for 373 year, and listed as #28 on Klaus Schmeh’s (the most-published cryptology author in the world) “The Top 50 unsolved encrypted messages.” The solution? Almost embarrassingly simple in hindsight.
I built github.com/dashaun/spring…, a zero-configuration #SpringBoot starter that brings AGENTS.md support to #SpringAI applications. #Java #AgentsMD #AI
Most annotation tools stop at drawing boxes. I’m building LabelOne — a local-first, open-source workspace for annotation, non-destructive image pipelines, synchronized views, and model feature inspection. Still early. Feedback welcome ↓ github.com/YounkHo/LabelO…
Introducing GLM-5.3-Flash - Leading capabilities at a highly competitive price - Natively multimodal with a 1M-token context window - A 320B-A18B model released under the MIT License - Previously previewed as Ox Alpha, running entirely on Chinese AI chips Blog: Show more
GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: huggingface.co/zai-org/GLM-5.3 Tech blog: z.ai/blog/glm-5.3
⚡Meet Qwen3.8-Flash, a multimodal MoE and an early preview of the Qwen4 architecture, now open-weight! The production version Qwen3.8-Flash will be available soon via QwenCloud API at just $ 0.16/1M input tokens and $ 0.47/1M output tokens. 125B parameters + 51B N-gram Show more
GLM-5.3-Flash can now be run locally! ✨ Run 3-bit on 128GB RAM via Unsloth GGUF. GLM-5.3-Flash (ox-alpha) rivals Claude Opus 4.8 on DeepSWE, coding & agentic benchmarks. Guide: unsloth.ai/docs/models/gl… GGUF: huggingface.co/unsloth/GLM-5.…
Introducing GLM-5.3-Flash - Leading capabilities at a highly competitive price - Natively multimodal with a 1M-token context window - A 320B-A18B model released under the MIT License - Previously previewed as Ox Alpha, running entirely on Chinese AI chips Blog:
GLM-5.3 can now be run locally! The 2-bit model retains ~81% accuracy after we shrunk it from 1.51TB to 239GB (-83% size). Run on a 256GB Mac or RAM/VRAM setups. GLM-5.3 is the strongest open model to date. Guide: unsloth.ai/docs/models/gl… GGUF: huggingface.co/unsloth/GLM-5.…
GLM-5.3 is now open-weight. Our most capable model for agentic coding and cyber defense is now available to download, run, and customize. Weights: huggingface.co/zai-org/GLM-5.3 Tech blog: z.ai/blog/glm-5.3
🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Cordis meta-framework, DeepSeek Harness is an agent harness built around one Show more
Glad that others are finding Amplio helpful. It is our main harness for autonomous long research runs spanning days to weeks. It is heavily used by our team for automated AI research / RSI on top of the Simply codebase. Open sourced at: github.com/google-deepmin…
Harness matters. We test Gemini 3.1 Pro with amplio, a harness for robust and long-horizon runs and super fit for auto-research. We then get much better results than with Gemini CLI. Harness: github.com/google-deepmin… FrontierCS is an open-ended benchmark for Computer Science
We’ve open sourced the engineering layer behind Dot Reflex 14B. The repo includes the inference runtime, local API, trajectory schema, agent loop examples, tests, benchmark receipts and training provenance. GitHub: github.com/usedotai/dot-r… Model weights: Show more
つーことで、リリースった。 "HRFFA: High-Angle Robust Fast Face Alignment. Whole-head face alignment dataset & training pipeline." github.com/PINTO0309/High…
Made a stand-alone Hermes Agent plugin for BackSearch - a wayback machine-like SaaS for agents - adds two tools, and requires their API key. Let me know if you like it github.com/NousResearch/h…
🔍 Introducing BackSearch. LLMs are increasingly asked to predict the future, but a good backtest requires a snapshot of the internet at a point in time. BackSearch allows LLMs to search the web as it was on a particular date. It’s great for: 🔮 Forecasting and prediction
we did it, fam this is my very first physical product took it from concept to manufacture, all inside of the CAD app I built what a time to be alive!
my physical product design era: in shambles in my first design review, with the department of home interior (my wife), the prior design was rejected for being “too chonky” to appease the department, I’ve gone back to the drawing board. The all new Stackable Coaster: Slim Mode
My LLM cliché highlighter is up to 38 patterns now tools.simonwillison.net/llm-cliche-hig…
Smart pick-and-place with OMY, Gemini, and GraspNet! We integrated VLM reasoning, GraspNet 6-DoF pose synthesis, and Cyclo Control for collision-free manipulation on ROS 2. 🎬 Video: youtu.be/OxMBVGkFJHM 🤖 Code: github.com/ROBOTIS-GIT/cy… 📝 Docs: docs.robotis.com/docs/systems/o…
It's #NationalParkWeek & me and my friends visit atleast 2 of them every year (it's our yearly ritual) - so, I built this all in one parks app a while ago called 'summit scout' 🏞️ summitscout-five.vercel.app... for all the adventurers out there, the app visualizes everything about Show more
We’ve built exactly this :) AO lets you manage Claude Code, Codex, Cursor, and 26 coding-agent harnesses from one desktop app. aoagents.dev github.com/Untrivial-ai/a…
Just released a sandboxed AI harness. It's written purely in Zig and utilizes things such as landlock, eBPF, and namespaces to isolate tool calls for LLM's. It also utilizes WASM for plugins in an isolated process and fed back to the session over IPC. github.com/Midstall/chock
We've just released code+model for Surflo!🚀 Turn any number of photos into a 3D surface with Flow Matching. Our latent representation encodes the full scene and can be queried for downstream applications like 3D segmentation. 🔗anttwo.github.io/surflo/ 💻github.com/Anttwo/Surflo
What if you could turn any number of photos (3, 8, 15, or even 60) into one clean 3D surface (pts & mesh) with Flow Matching? Check out our new work, Surflo: Consistent 3D Surface Flow Model with Global State. 🧵 1/n 🔗anttwo.github.io/surflo/
We’re not just building @RomeAILab — we’re running our company on it. As an AI-native company, we use Rome every day to help us get work done. We’ve open-sourced some of the apps we use internally so you can see how we actually work with our agents. github.com/rome-os/rome-a…
Congrats to @Zai_org for another great release! It works best on Claude Code, which unfortunately doesn't support it natively. So I built an installer that enables GLM-5.3 to be selected inside Claude Code (living side by side Claude models). github.com/xhluca/claude-…
GitNexus MCP plus CLI in one package is how I want agents to talk to a repo github.com/abhigyanpatwar…
Need to blur a face? You don't need to upload the video. I built Redact, a Mac app that redacts faces + objects entirely on-device. @ultralytics YOLO26-seg + Neural Engine. No upload. No account. It also tracks objects between detections, so one missed frame doesn't leave an Show more
A fun experiment with Code Mode and Flue to make generative art. Instead of inventing a tool for every possible shape, the model can write JavaScript and run it in an isolated worker. repo: github.com/paperwing-dev/…
grok voice is fast. i built something that would make that speed visible. my open source engine draws what you speak, while you're still speaking. numbers become live charts. processes become diagrams. it cannot hallucinate a number, it only plots what was actually said. right Show more
We open-sourced all components of the HTD whole-body controller (WBC): github.com/chrisyrniu/Isa… -Play our WBC on your browser: humanoid-touch-dream.github.io/wbc_mujoco/liv… -Main HTD codebase entry: github.com/chrisyrniu/hum… (data collection and HTD policy soon!) Highlighted features: -Lightweight Show more
A touch-aware humanoid manipulation policy that cleans the lab for you🧹🧪 Introducing Humanoid Touch Dream: a real-world system for dexterous, contact-rich humanoid loco-manipulation. Our key idea is simple: the policy predicts future hand forces and tactile latents alongside
We released DROID+, a new and improved DROID control stack from SRL. We believe robots (and hardware) should be easy to setup and easy to use, and so we updated the DROID stack to be the same. @HugoHadfield1 github.com/nvlabs/droidpl… Here's pi0.5 doing our "hello world" task:
You might've heard of @OpenAI, @OpenRouter, @opencode We're excited to launch @TryOpenTag : the model-agnostic Claude Tag. If you've seen Claude Tag, you know the interface. Tag an AI in Slack, hand it a task, get the result in the thread. We think that's the right interface. Show more
🚀 PointDiT is now completely open-source! github.com/google-researc… We’ve officially released the code, pre-trained models, and full training/eval pipeline. Hope this helps the community! ✨ Show more
Does 3D reconstruction have to be complex? We answer this question with PointDiT (#ICML2026): a minimalist pixel-space Diffusion Transformer without bells and whistles. We show that a plain ViT can estimate dense 3D point maps by operating directly on raw patches. No hybrid
God’s Eye View V1 is open source. It feels like a spy simulator in your browser -- except the underlying data is real, and you can talk to the planet too. Track planes, ships, satellites, traffic cameras, and critical infrastructure. Use voice mode to ask what’s happening, Show more
Have been using qwen3.8:27b for my local hermes agent. Had some issues with compaction and thinking modes. This plugin I made fixed all the issues I was having, maybe others will benefit too github.com/benthecarman/h…
sharing my first blog! a meta harness for self-improving AI chip design… check it out if u wanna see how RL could work in chip design task, or how an ASIC design can become a hillclimbing task it introduced ASIC vs. GPU, RL Env setup, and the best generated design can run Show more
open models are getting really good, but using them still feels like something made for developers. i think everyone should have access to them. so i built beacon - a ios app for downloading and chatting with open models directly on your device. models come from hugging face, Show more
Perfect timing with the “malleable computer for the age of agents” vision. I built a thin harness that lets an agent fully control an Omarchy desktop (Hyprland, AT-SPI, real browser, etc.): github.com/fabiopauli/oma…
too many people DMd me about this so i open sourced it if you read a lot of stuff and use hermes and wanna super charge it, go for it github.com/ianlapham/supe…
Today we're launching Proliferate, open source Codex for working with any agent- here's how we use Proliferate to build Proliferate.
got tired of clicking through menus just to see how much Codex I’ve burned through So I built a tiny macOS menu bar app to keep the usage in check github.com/gameofbitcoins… Tibo @thsottiaux - maybe Codex should just ship this by default?
I built a harness that gives every AI bot its own directory, its own browser profile, its own cookie jar, its own environment. -> github.com/Archive228/cub… Bot asks for a service → a grant names it, or it's denied. Grant expires → the door closes on its own. Handoff → only the Show more
I made a quick performant dashboard for multi DGX Spark setups, view throughput of your running inference containers, link throughput, GPU processes, and more. Suggestions and contributions are welcome! github.com/metaspartan/sp…
Frustrated your coding agent keeps ignoring your rules? What if every rule could take on a life of its own? I built RulesAsPrograms which turns each rule into a function that checks agent's work. Instead of hoping agent follows rules, rules check agent github.com/programasweigh…
👋📊 I built a spreadsheet webcam emulator app in @GoogleAIStudio (using Gemini 3.7 Flash), so you can play around with this yourself! Also allows you to export images and CSVs: spreadsheet-webcam.ai.studio
we are roughly 48 hours away from someone streaming the entire @wnba playoffs inside =INDEX(MATCH())
I open sourced Slopdar. Now I want to see what happens when other builders start messing with it. Add a weird AI pattern. Improve a detector. Break something. Fix something. Build something on top of it. The repo is completely open 👇 github.com/Slopdar/slopdar
I made the fastest inference stack for Qwen3.8-27b on a RTX 3090. Up to 195k context, 82 tps single request, 672 tps peak 64 concurrent. github.com/syv-ai/qwen38-…
Day-one Qwen3.8-27B. NVFP4, 262K context, one DGX Spark, one command. Open-sourced the full reproducible deploy. Still Tuning IT ! github.com/tonyd2wild/Qwe…
A 27B model on two consumer 3090s at ~75-85 tok/s w/ NVLINK Qwen3.8-27B FP8 with tool calling, reasoning, vision, and MTP speculative decoding. NVLink or PCIe. Full deploy open-sourced. github.com/tonyd2wild/Qwe…
I made PARS//Poole Adaptive Reasoning Stack into a skill! Generate exact binaries, research, and audit with PARS. Currently only tested and used with Codex 5.6 Sol. Get the skill.md here: github.com/rookepoole/PARS Support my research and builds here: buymeacoffee.com/rookepoole
Agreed, been working on this. Agents re-learning who you are every session is the real waste. Mid-session compression alone cut 61% of token usage in our case. open-sourced the whole thing 👉 github.com/TencentCloud/T…
I built an open Agent Memory for Cloudflare. CF's version is still private beta. I needed profiles, extract, and recall now, so I shipped my own on Workers + Durable Objects. Cheap Luna for extract and answers. One-click deploy. MIT. github.com/thisuxhq/agent…
👀 Treg is open sourced here: github.com/superdesigndev… If you are vendor feel free to PR to list
Introducing OpenRouter for tools ⚒️ Old world: SaaS bundles priced for humans. $139/mo, and you don't even know what's in the box. New world: agents don't care about vendors. They want the best API for the task - and to pay for the result. - 2,600 agent-friendly tools:
Happy to announce that the game I've been building for the past 2+ months now has a playable demo! If you're at your computer and would like to give the game a try, you can play it at dannylimanseta.itch.io/cook-the-dunge… I'd love to hear your feedback and suggestions in the itchio comments Show more
Make your fingers dancing. a little toy I made , just 85 lines made by claude code and me . This is actually part of the video data cleaning and filtering process for Physical AI data collection. github.com/fat1/handMedia…
One week ago, @sumnuz and I built Grand Theft Austerity at @elitepax’s AI Game Jam. One week and about 6.68B tokens later the open-world Bucharest version is live. 🇷🇴🌍 🎮 Play: gtausterity.vercel.app 💻 Repo: github.com/alexdevmotion/…
Loved making Grand Theft Austerity with @sumnuz at AI Game Jam - a Bucharest satire of servers, broadcasts and battered Dacias. Thanks @elitepax for organizing, and congratulations to everyone who built such amazing games! 🎮
I built a Resume Classifier. Here is the to check it out classr.streamlit.app Github Repo :github.com/vee234o/classr More info here :linkedin.com/posts/oreoluwa…
I just shipped Nado🎉 You type an order like texting a customer, get back an AI parsed invoice, a payment link, and an auto-generated receipt. I built this for the @apiconflagos X @monnify Developer Challenge #APIConfXMonnify #DeveloperChallenge 🔗: usenado.vercel.app
Claude Cowork is impressive. But its limits can be brutal. Computer-use tasks may consume far more usage than normal chats, and even expensive subscriptions have caps. So we built and open-sourced an alternative. Backed by Y Combinator. github.com/coasty-ai/open…
Introducing Autoprover. AI’s newest math meme: give a model an open problem and keep replying “keep going.” Sometimes it actually works - so I built a harness that automates that flow. Just on steroids. github.com/guyz/autoprover Show more
here's how we did it and the open sourced @HyperFrames_ code in thread 🧵 i did a quick screen recording as a quick explainer our open source framework & motion graphics skills: github.com/heygen-com/hyp…
we've eval-ing K3 and we were legit shocked K3 is incredibly good at @HyperFrames_ , winning over GPT-5.6 and Fable 5 at a few of our benchmarks it recreated motion graphics from @leomeethewoo with minimal prompting (i didn't further tune the animations to show raw output)
So I built something on the cognee cloud for the wemakedevs hacakthon continuum-sigma-three.vercel.app The healthcare system forgets. Continuum remembers A cross-institutional patient memory layer that turns fragmented medical records into one intelligent, timeline using Cognee's memory
btw, I built this. Any feedback or suggestions are appreciated. github.com/Da7-Tech/mind
Just 6 classes later: I built my own ChatGPT-Chaigpt. 🚀 Grateful to @surajtwt_ for the guidance and to the #GenAI Cohort at ChaiCode for making this possible. Thanks @Hiteshdotcom @ChaiCodeHQ @piyushgarg_dev @nirudhuuu @devwithjay @yntpdotme Link-chai-gpt-alpha.vercel.app
Here're some of the great stuff people are currently building ✨ You're welcome to add more in the comments. @SamJWasserman 10 apps/tools to help with AI native filmmaking github.com/wassermanprodu… @sfxnz An open sourced tool that allows anyone to serve and run evals for local Show more
if you want to run this yourself, the developer (slvDev) open-sourced the entire project here: github.com/slvDev/esp32-ai
projects i built in the past few months: - z-studio.vercel.app - lattice-doc.vercel.app - bounty-stack.vercel.app - pixel8-ui.vercel.app - who-wants-to-be-a-king.vercel.app - github.com/ANAS727189/Tru… - github.com/ANAS727189/San… - anas727189.itch.io/bolt + a few more on my GitHub👇 Show more
Didn’t expect the last post to get this much love. Thank youuu Everyone! ❤️❤️ More projects i built: Svelte AI Elements: svelte-ai-elements.vercel.app Svelte 4 Animations: old animation-svelte.vercel.app Svelte Form Builder: svelte-form-builder.vercel.app
wtf this is Crazyyy sv-animations.vercel.app sv-table.vercel.app sv-blocks.vercel.app sv-efferd.pages.dev sv-matrix.vercel.app sv-particles.vercel.app sv-agentation.com
🔥 HumanCLAW is now fully open-sourced! github.com/Human-CLAW/Hum…
🦾HumanCLAW: Can Vision-Language Models Act Through a Body? A VLM can spot the sofa in a second. But can it get a body there and sit down? We call it Action Intelligence, and our work decouples it from motor control so it can finally be measured. Today’s best models turn out to
During early access to GLM-5.3, I built this website with 25 separate small biology applications, each of which is quite useful for day-to-day planning and performing experiments. It only took a couple of hours and a few edits. That’s pretty remarkable! benchkit-sigma.vercel.app
I made a tool that reason around a question without context window limits, and visualize the knowledge graph github.com/punnerud/Local…
I open sourced this idea AND called it “mind virus” like a year ago btw: github.com/dadukhankevin/…
I built my own AI coding tool that's probably better then Claude Code using Gemini 3.7 as the default model. Try it here using your own Google AI Studio API key. MinovativeMind.dev I want someone to prove me wrong that my tool is not better then Claude Code at coding.
Hmm Already doing the extreme low-end version of this: DeepSeek-V4-Flash-0731 on a 6yo laptop with 11 GB usable. One layer lap is 1.649 GiB so LRU under that size is pure 0% hits. Engine is public. github.com/Heman10x-NGU/e… Maybe i ll ask codex to improve my inference engine as Show more
I wanted to understand how @PyTorch works. So I did the obvious thing: I built it from scratch in C. github.com/thevoxium/bare…
I just open sourced a minimal chatbot template. This brings together everything we've been doing the past weeks: new components, shadcn/typeset, shadcn/react, tool calling, HITL with questionnaire, message parts. Deploys one click to @vercel AI Gateway. github.com/shadcn-ui/chat…
really cool work! i built an oh-my-pi extension inspired by this workflow - github.com/wolfiesch/omp-…
Scaling self-verification with DeepSeek V4 Flash beats Claude Fable 5 on Terminal-Bench 2.1, while being 11x cheaper 💰 As open-source models become more capable, they can now generate large numbers of high-quality candidate solutions and verify their own outputs at very low
I run Claude Code, Cursor, Codex, and OpenCode in parallel. They had no idea the others existed, even in separate worktrees, they'd redo or contradict each other's work. So I built a shared "brain" for them to plug into: a local app, an MCP, and a CLI. It drives the CLIs you Show more
I’ve been experimenting with bringing on-device LLMs to Flutter using Qualcomm’s QNN - and the results are pretty interesting. I built Flutter-QNN, a Flutter app that runs LLMs directly on Snapdragon devices, using the Gemma Flutter (flutter_gemma) package and QNN to take Show more
I built something for this problem. It adds X, YouTube, RSS, Bluesky, Mastodon and other importers. It has a web and MCP interface too - so you and your agents can search it. github.com/darron/dbrain
Guess what AI Model built this game? Local Qwen 3.8 27B (NVFP4) via Hermes wrote a single-file HTML canvas Breakout. This model is straight fire 🔥
Cybercrime is projected to cost the global economy $10.5 trillion every year. AI has made every company, everywhere, vulnerable The problem isn’t a lack of security tools. It’s that they were never designed to defend against AI threats We built Trident (YCS26) to change that, Show more
We launched AI Search last year — and Shoply is already ranked #2 for “AI Search” on the Shopify App Store. 🚀 See the ranking: apps.shopify.com/search?q=AI%20… Try Shoply AI on Shopify: apps.shopify.com/shopping-assis… #Shopify #AISearch
🇮🇳 Happy Independence Day to all of us! This Independence Day, let’s make a habit of being informed investors: making decisions with better information, research, and awareness. We’ve also just launched our Research AI, built to help you research companies, understand Show more
FRIDAY PROJECT READY : AGENTZERO 🌿 @CNPYNetwork asked, "What are you building this week ?", so this time I'm coming with a screenshot. AgentZero, which I built on the Canopy testnet, is my own chain where I'm testing the AI native Layer-1 concept. I designed AgentZero Show more
Happy Friday builders 🌿 We want to see what you’re creating on Canopy. Screenshot it, drop it below. Ugly early versions welcome (and encouraged!) Best one this week gets a repost.
Hi, I’m a Software Engineer 👋🏾 I build scalable web products with Next.js, React, TypeScript, TanStack & Node.js. Just launched Jargons 🦉, an open-source AI code review & auto-fix agent. jargons.run Open to paid roles & gigs. devbio.co/jay
It’s weekend 🎉 Quote this tweet with what you do, shamelessly sell yourself, tell us what you do 💗 If you’re hiring, Quote with your vacancies, let’s get people hired 🎉
Your coding agent can see the image. You can't. Claude Code, Codex, and Hermes all run a full-screen TUI that owns the terminal. A tool writes a chart, and you get "Read image (42 KB)". So I open-sourced termpeek. It renders where the TUI never repaints, in a tmux pane Show more
I wanted ChatGPT to operate the Mac where my work lives. So I open-sourced Mac Developer Bridge: shell, files, real PTYs, background jobs + read-only Codex history. MIT. No model calls. Very much not sandboxed. github.com/alexanderradah…
After 10+ years working with SMBs, I’ve seen how hard it is to get visibility into customer feedback, conversations and sales. I built an AI app to solve this. Just launched on Product Hunt. Would love your feedback 👇 BRIKS: Chat with your business producthunt.com/products/briks…
Scrimba just launched a mind-blowing new learning tool: Explain. You ask any question or text and it turns that into a video explainer instantly 🤯 There's also a Chrome extension available. And you can even use it via ChatGPT/Codex and Claude plugins. They're now LIVE on Show more
- wrapped up p2p side of @b3t_protocol - launchpad didn’t exist for how i would want to launch a token, so i built one - had some random shit i wanted to brush up on and study. built claude wrapper/pipeline to turn it into textbook style reading on a blog - don’t wanna Show more
Can AI create a luxury campaign that actually feels like a commercial? I built this fictional fragrance concept from two references: one product, one character. 15 seconds. One consistent visual story. NOIR ÉLAN — *The Last Detail.* 🎬 ✨ Created with Grok Imagine @grok
The City That Recognized Her 30 million faces in the city. Tonight, every screen showed hers. And then the drones started looking. ✨ Created with GPT Image 2.0 @ChatGPT Prompt below 👇 Create a cinematic dark cyberpunk scene at midnight in a gigantic futuristic megacity
The most expensive minute in mobile security is the time between your app shipping and an attacker reading it. Trawlpost eliminates it — AI agents crawl the App Store and Play Store 24/7, dropping MASVS and GDPR digests on every new release the moment it lands. Live soon.
A 19-year-old Japanese student built a trading bot with Claude Code in 2 days. Used his iPad as a second monitor. First night: $6,732 profit. Starting capital: $68. Total profit so far: $750,000. Here's how it works: The bot scans over 50 markets simultaneously. Syncs live Show more
$0 a month on cloud AI. $2,300 once on a box that pays for itself in a quarter. Here is the whole build. The GPUs: 2 used RTX 3090s. 24GB each, 48GB total. That runs a 70B model quantized, on your desk, offline. The base: a Chinese X99 dual-Xeon board, $180 shipped. 2 used Show more
Gemini just hit 1 billion users. AI is not the hard part anymore. The hard part is still the same: getting the thing in your head to exist, live, for people to use. That's the gap we built Vibely (vibely.sh) for. What haven't you shipped yet? #buildinpublic
We just launched the AI Market Dashboard on The Groton Study. AI moves incredibly fast, and keeping up with everything happening across the industry can be difficult. We built this dashboard to make that easier. It continuously tracks and organizes: → Biggest AI stories → Show more
We built this AI Cowork harness and tested heavily with open weight model. We discovered that DeepSeek V4 Flash work extremely with this harness. But recently Meta just launched muse-glimmer-30b, they claim that this model was trained to become agentic out of the box (know how Show more
The most useful thing I learned building two cancer tools as a practicing oncologist: I was wrong about which one people would want. I built the AI tumor board first, because that is my daily frustration. Paste a de-identified case, an AI panel reasons through it, it matches Show more
Yea, I built an inline prompt optimizer that just fires off a forked thread and returns improved prompt with diff view component. Some stability issues with overall tool, but imagine those will get ironed out soon.
Built this cool interactive website for human heart anatomy using Gemini 3.7 Flash in AI Studio. You can explore the heart in 3D and learn about its different parts.
Magnific just shipped Auto Layers 2.0, and it quietly fixes the most annoying part of AI image work. It splits any image into editable layers. Subject, text, logo, background, all separated automatically. You upload or generate an image, get every layer on its own, then move, Show more
This week, I built an AI agent designed to help application owners and software packaging teams automatically generate software installation instructions reducing the manual effort required to create, document, and standardize installation procedures.
Apps and App Builder are live in Snaptrude. Turn your firm's design workflows into software. Describe the workflow, define the rules, get an App that runs on your building model. Head to snaptrude.com to try it out
Made this 3d Japanese tofu drift game today, fully built on Crayon Pro, access going live soon 👀 you can just build beautiful and fun games #indiegame #gamedev #threejs
Introducing GLM-5.3: Built to Code. Ready for Cyber Defense. - Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model - A major leap in cybersecurity, setting a new standard among open models Tech Blog: z.ai/blog/glm-5.3
launching openrec.co, a free and open source granola alternative it's the 2137th open source granola alternative, except this one works and looks good. i made this out of curiosity, following the trend of replacing saas with free software, wanted to see how hard Show more
Here’s another app I’m releasing: LedgerPDF, a free and open-source workpaper binder for tax and accounting professionals working alongside AI agents. I built it for the same reason I built LedgerTB: I want to see and review what an agent does. If an agent prepares my Show more
1/ I built an autonomous rover in my mom's basement completely from scratch. chassis, drivetrain, PCBs, STM32 firmware, controls, inference, all of it. In this demonstration, it’s controlled by a neural network, trained to push objects into a goal zone (like Ronaldo on wheels). Show more
Its two days into fungi week here is what has been released if you missed it Agent tools for any SDK that live refresh, can be searched and described and have an attach/dettach state machine. npmjs.com/package/@shiit… npmjs.com/package/@shiit… A Durable Agent mailer and the Oban Show more
10 fresh indie builds just went live on InventList: Procright · Anomaly Ai · Bloom · Popform · Max · Bon Split · Archmaster · Flexify Suite · Kage · Roman Names Dev tools, AI agents & SaaS from makers shipping in public → inventlist.com/signals/week/2…
Just launched today pgbot.dev 🤖🐘 Postgres intelligence for agents & apps. Free. Open source. Written in Go. Nobody wants to stare at PostgreSQL dashboards all day. So I built a bot that does it for you. → finds what’s wrong → explains what changed → tells Show more
Launched today pgrun.dev metrics 📈 What PostgreSQL metrics should you monitor in production? → CPU + load → Memory → Disk usage + I/O → Network traffic → Active / total connections → Cache hit ratio → Query throughput → Deadlocks → Database size →
Lol - fully shipped this agent with persistent memory in less than a day. Cooking with gasoline. AI is insane
Meetily is a 100% open source meeting assistant built on Rust and Tauri that transcribes 4x faster with Whisper and Parakeet, all running locally with no cloud dependency. → Runs summaries locally via Ollama or custom endpoints like Claude and Groq → Captures mic and system Show more
Episode 44 Continuing to track early signals across AI, Robotics & Bio — another batch of teams building across robotics, aerospace, AI agents, enterprise software, and frontier technology. Show more
corrected W32 card: the earlier image showed only a subset. This attachment shows all 20 projects from Aug 3-Aug 9, 2026, in manifest order. 1. ds4: DeepSeek V4 and GLM inference across Metal, CUDA, and ROCm. github.com/antirez/ds4 2. Computer (Cloudflare): Durable Object Show more
corrected W32 card: the earlier image showed only a subset. This attachment shows all 20 projects from Aug 3-Aug 9, 2026, in manifest order. 1. ds4: DeepSeek V4 and GLM inference across Metal, CUDA, and ROCm. github.com/antirez/ds4 2. Computer (Cloudflare): Durable Object Show more
corrected W32 card: the earlier image showed only a subset. This attachment shows all 20 projects from Aug 3-Aug 9, 2026, in manifest order. 1. ds4: DeepSeek V4 and GLM inference across Metal, CUDA, and ROCm. github.com/antirez/ds4 2. Computer (Cloudflare): Durable Object Show more
corrected W32 card: the earlier image showed only a subset. This attachment shows all 20 projects from Aug 3-Aug 9, 2026, in manifest order. 1. ds4: DeepSeek V4 and GLM inference across Metal, CUDA, and ROCm. github.com/antirez/ds4 2. Computer (Cloudflare): Durable Object Show more
corrected W32 card: the earlier image showed only a subset. This attachment shows all 20 projects from Aug 3-Aug 9, 2026, in manifest order. 1. ds4: DeepSeek V4 and GLM inference across Metal, CUDA, and ROCm. github.com/antirez/ds4 2. Computer (Cloudflare): Durable Object Show more
PentesterFlow: CLI agent for AI-assisted penetration testing PentesterFlow is an open-source CLI agent that follows a plan → act → observe → verify → report cycle. It supports local and hosted LLMs through Ollama, LM Studio, Gemini, Groq, and OpenAI-compatible APIs. The Show more
🇺🇸This might be one of the wildest things to come out of Grokathon. A team connected Grok Voice to brain-sensing tech and built an app that lets someone talk… without ever opening their mouth. Imagine giving a real voice to people who physically can’t speak. Absolutely Show more
Grok Build is way more powerful than most people realize. SpaceXAI has quietly shipped a ridiculous number of features in a very short time that most people don’t even know exist. Here are the biggest ones you should know about
Why Robotics Is Entering Its “Open Source” Moment Linux changed software by proving that a powerful system could be built through contributions from a global community rather than controlled entirely by one organization. Wikipedia showed that knowledge could grow through Show more
Wouldn’t it be amazing if an e-commerce app could actually show you exactly how that shirt will look on you before you buy it? 🤯 That’s what I built using Claude Opus 5. A real fabric and size aware virtual try on experience where you submit your measurements and photos, Show more
The original experiment was 11 Claude Code agents. The OSS repo is now: → 18 reasoning personas → Claude Code + Codex + Gemini CLI + OpenCode → multi-provider routing → Full / Quick / Duo deliberation modes → confidence-weighted verdicts → ~3.8k GitHub stars Turns out Show more
No time to screen a flood of applicants for a role at our family office, so instead of hiring a recruiter, I had Claude Code build a program that interviews candidates over chat, scores their fluency and how they think, and flags likely AI-assisted answers. It's live and Show more
Eliza + elizaOS Welcome to the new interaction paradigm. Notes, calendar, browser, messaging, wallet, social media and everything else you use your phone for. Completely driven by chat and voice. Eliza comes as an app on mobile and desktop and as a full operating system built Show more
Someone vibe coded an AI interior tool that redesigns your room using furniture you can actually buy, with every piece linked to a real product and a full shopping list in the sidebar. x.com/om_patel5/stat…
An AI assistant just got its own phone number and email address. It makes your calls, waits on hold, and books the follow-ups. And it is one of 7 things that dropped this week most people slept on. Me and @AndrewWarner broke them all down in the video below. Here's what you Show more
An AI assistant just got its own phone number and email address. It makes your calls, waits on hold, and books the follow-ups. And it is one of 7 things that dropped this week most people slept on. Me and @AndrewWarner broke them all down in the video below. Here's what you Show more
An AI assistant just got its own phone number and email address. It makes your calls, waits on hold, and books the follow-ups. And it is one of 7 things that dropped this week most people slept on. Me and @AndrewWarner broke them all down in the video below. Here's what you Show more
Meet the Gemma Translator! A fully offline device powered by Gemma 4 E2B built with @Antigravity. Running entirely on a Raspberry Pi 5 with a connected microphone and speaker, this highly portable prototype is housed inside a custom, 3D-printed case.
was just reflecting on everything i've done this year. for a moment, i genuinely felt like... "have i even built anything?" then i opened my GitHub and scrolled through my X profile. turns out, i've shipped more than i gave myself credit for. over the last 3 months, here's Show more
was just reflecting on everything i've done this year. for a moment, i genuinely felt like... "have i even built anything?" then i opened my GitHub and scrolled through my X profile. turns out, i've shipped more than i gave myself credit for. over the last 3 months, here's Show more
was just reflecting on everything i've done this year. for a moment, i genuinely felt like... "have i even built anything?" then i opened my GitHub and scrolled through my X profile. turns out, i've shipped more than i gave myself credit for. over the last 3 months, here's Show more
You can now run a 2.78 trillion parameter model on a CPU with 8.24 GB of RAM 🤯 kimi-k3-in-c is a 176KB pure C99 engine that streams Kimi K3’s experts from disk instead of loading them into memory. no GPU. no CUDA. no framework. 100% Open Source.
Prime Agent is a general-purpose coding harness On ARC-AGI-3, it scores 95.5%, surpassing the human-expert baseline, but the gain is not benchmark-specific. We see major improvements across models when compared to their proprietary harnesses:
Today we are releasing Cloudflare OS, a chatbot with connectors, just like every other tech company is doing. Except actually, it's different. This is a remake of Sandstorm[.]io, my startup from 10 years ago, except this time built on Cloudflare Workers (the platform I've spent Show more
What shipped so far: - A voice AI product with live coaching reports (talk ratio, objection handling, buyer verdict) - Top 5 Product of the Day on Product Hunt - First paying client at $20K - Currently: 20 free sales tools in 20 days, build in public
introducing anydoc now your agents get 100x faster local parsing for pdf, docx, pptx & 10 more formats - sub-5ms md conversion - 500 docx files in 1.7s - top quality across all 13 formats - rust-based - open source already powering @firecrawl /parse github.com/firecrawl/anyd…
Today we’re partnering with AWS to launch Superblocks 3.0: the secure way for employees to vibe code production enterprise software. In a single prompt, Superblocks can replace million dollar SaaS, while IT & Security stay in control. OpenAI and Anthropic are releasing new Show more
built forge for this exact reason, makes ai agents ship full projects from a prompt, forgee.xyz if u wanna check it outforgee.xyz if u wanna check it outforgee.xyz if u wanna check it outforgee.xyz if u wanna check it Show more
btw this robot runs on a fully open source AI framework called Xiaozhi it turns a $8 ESP32 chip into a fully autonomous voice assistant NO expensive cloud servers or high latency wake-word detection and audio streaming happen locally on the chip: > requires only an ESP32-S3, Show more
meet WALL-E no, not the one from the cartoon :) this is a fully 3D-printed, autonomous desk companion running on an ESP32. it uses a local AI voice assistant to process commands in real-time and control its own motors building intelligent robotics just dropped to zero cost...
The ai agent control pattern is getting more validated, the reason we built prismor and used by 8k devs monthly github.com/PrismorSec/pri…
Another week, another shoutout to the amazing makers behind these impressive products I discovered on 𝕏: QM (qm.ycombinator.com) by @ycombinator workbench.md by @mattshumer_ migma.ai by @TheDreadCEO exe.dev by @davidcrawshaw agentOS Show more
Another week, another shoutout to the amazing makers behind these impressive products I discovered on 𝕏: QM (qm.ycombinator.com) by @ycombinator workbench.md by @mattshumer_ migma.ai by @TheDreadCEO exe.dev by @davidcrawshaw agentOS Show more
I love this opensource project! Someone just put a 28.9M parameter AI model on an ESP32-S3. A chip that costs only a few dollars. No cloud. No API. No internet. It runs locally at ~9.5 tokens/s and can generate stories on a tiny screen. This changes the way we think about AI Show more
they gave frontier models a 1 billion token context window: you can now run Fable [1B], Opus [100M], GPT 5.6 [500M] across your entire codebase - it's a graph based tool that can traverse databases with 400M+ of nodes in ms - only one prompt to install - & it's open source Show more
So what are we actually building at Axis Robotics? Four products, one compounding loop: 1. Task Generation Engine (live) Give it a prompt like: “A study desk with a pour-over kettle, a hand-crank coffee grinder, a coffee mug, a notebook, and a pen.” The engine turns that Show more
At Axis Robotics, our vision is to build a compounding data engine—one that connects large-scale pretraining data, corrective post-training data, model deployment, and failure feedback in a continuously improving loop. Over the next 6–12 months, we will advance this vision
this creator just built a zero-delay auto-aim system on an $8 microcontroller he deployed a custom local AI algorithm on a cheap ESP-32 to track human movement with absolute 0-pixel accuracy. the system completely eliminates standard computation delay. it processes the bounding Show more
finally! local AI is accessible to everyone. we replaced our massive smart home server with a single $8 ESP32-S3 chip processing voice commands entirely on-device. now anyone can build this. here is the full guide: x.com/i/article/2081…
One of the coolest open-source projects I have seen is text-to-cad. Describe a part in plain English, and AI helps generate production-ready CAD artifacts. It supports: - STEP, STL, 3MF, DXF & GLB export - URDF, SRDF & SDF generation for robotics - CAD model editing through AI Show more
text-to-cad just crossed 10,000 stars you can use it to generate: - STEP files - URDF, SDF and other sim artifacts - 3D mesh files like STL, 3MF and GLB - gcode for 3D printing - DFM checks for popular services like sendcutsend 100% free and open source ✌️
An example of an AI-assisted project — Slotbound. The developer said on Reddit that he had no prior game development experience, but had always wanted to make games. 8 months ago, he decided to try bringing the idea to life with the help of AI. The demo has an 80%+ positive Show more
Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document, send a slack message, or update a calendar entry. Ask it to prepare a customer brief, untangle your calendar, draft a report, or triage a Show more
Introducing Muse, a personal agent that gets things done for you, powered by Muse Spark 1.3. Get an inside look at how we built Muse and what it can do for you: introducing.muse.ai
Introducing Muse, your personal AI agent from Meta that gets things done across every part of life. Download the Muse app and get started: Muse.ai