{"exhaustive":{"nbHits":true,"typo":true},"exhaustiveNbHits":true,"exhaustiveTypo":true,"hits":[{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"lwhsiao"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"<em>Oh-my-pi</em>: A coding agent with the IDE wired in"},"url":{"matchLevel":"none","matchedWords":[],"value":"https://omp.sh/"}},"_tags":["story","author_lwhsiao","story_48994611"],"author":"lwhsiao","children":[48995761,48999173,48999389,48999727,49000430,49007232,49022642],"created_at":"2026-07-21T16:32:45Z","created_at_i":1784651565,"num_comments":15,"objectID":"48994611","points":42,"story_id":48994611,"title":"Oh-my-pi: A coding agent with the IDE wired in","updated_at":"2026-09-09T21:27:22Z","url":"https://omp.sh/"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"maherbeg"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"<em>Oh-my-pi</em> \u2013 The Harness Playbook"},"url":{"matchLevel":"none","matchedWords":[],"value":"https://stencil.so/blog/harness-playbook"}},"_tags":["story","author_maherbeg","story_49539332"],"author":"maherbeg","children":[49542002,49542326,49545171,49545358,49582994],"created_at":"2026-09-02T17:13:30Z","created_at_i":1788369210,"num_comments":12,"objectID":"49539332","points":35,"story_id":49539332,"title":"Oh-my-pi \u2013 The Harness Playbook","updated_at":"2026-10-05T12:38:44Z","url":"https://stencil.so/blog/harness-playbook"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"flashblaze"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"Having fun with <em>oh my pi</em>, DeepSeek-V4-Flash, GPT-5.6 Luna and Antigravity CLI"},"url":{"matchLevel":"none","matchedWords":[],"value":"https://flashblaze.xyz/posts/having-fun-with-omp-deepseek-luna-and-agy/"}},"_tags":["story","author_flashblaze","story_49145463"],"author":"flashblaze","children":[49147584,49147889,49148453,49236380],"created_at":"2026-08-02T15:23:59Z","created_at_i":1785684239,"num_comments":9,"objectID":"49145463","points":25,"story_id":49145463,"title":"Having fun with oh my pi, DeepSeek-V4-Flash, GPT-5.6 Luna and Antigravity CLI","updated_at":"2026-10-05T12:49:57Z","url":"https://flashblaze.xyz/posts/having-fun-with-omp-deepseek-luna-and-agy/"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"carrja99"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"Current Agentic Development Environment: Orca and <em>Oh My Pi</em>"},"url":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"https://james-carr.org/posts/2026-09-30-current-agentic-development-setup-orca-<em>oh-my-pi</em>/"}},"_tags":["story","author_carrja99","story_49916232"],"author":"carrja99","children":[49916233,49951993],"created_at":"2026-10-01T00:27:00Z","created_at_i":1790814420,"num_comments":2,"objectID":"49916232","points":6,"story_id":49916232,"title":"Current Agentic Development Environment: Orca and Oh My Pi","updated_at":"2026-10-06T18:16:33Z","url":"https://james-carr.org/posts/2026-09-30-current-agentic-development-setup-orca-oh-my-pi/"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"dougcalobrisi"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"Tuning a Local Coding Agent: <em>Oh My Pi</em> and Qwen3.8-27B on Two RTX 3090s"},"url":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"https://doug.sh/posts/tuning-a-local-coding-agent-<em>oh-my-pi</em>/"}},"_tags":["story","author_dougcalobrisi","story_49711316"],"author":"dougcalobrisi","children":[49711345],"created_at":"2026-09-15T12:09:21Z","created_at_i":1789474161,"num_comments":1,"objectID":"49711316","points":5,"story_id":49711316,"title":"Tuning a Local Coding Agent: Oh My Pi and Qwen3.8-27B on Two RTX 3090s","updated_at":"2026-10-05T12:57:30Z","url":"https://doug.sh/posts/tuning-a-local-coding-agent-oh-my-pi/"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"dougcalobrisi"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"Custom Models in <em>Oh My Pi</em>: vLLM, Llama.cpp, SGLang and More"},"url":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"https://doug.sh/posts/<em>oh-my-pi</em>-custom-models/"}},"_tags":["story","author_dougcalobrisi","story_49852029"],"author":"dougcalobrisi","created_at":"2026-09-26T00:53:29Z","created_at_i":1790384009,"num_comments":0,"objectID":"49852029","points":2,"story_id":49852029,"title":"Custom Models in Oh My Pi: vLLM, Llama.cpp, SGLang and More","updated_at":"2026-09-26T01:36:12Z","url":"https://doug.sh/posts/oh-my-pi-custom-models/"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"_superposition_"},"title":{"fullyHighlighted":true,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"<em>Oh My Pi</em>"},"url":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"https://github.com/can1357/<em>oh-my-pi</em>"}},"_tags":["story","author__superposition_","story_48850500"],"author":"_superposition_","created_at":"2026-07-09T18:36:43Z","created_at_i":1783622203,"num_comments":0,"objectID":"48850500","points":2,"story_id":48850500,"title":"Oh My Pi","updated_at":"2026-07-09T18:48:31Z","url":"https://github.com/can1357/oh-my-pi"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"hantusk"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"Pi vs. OMP (<em>Oh My Pi</em> Agent)"},"url":{"matchLevel":"none","matchedWords":[],"value":"https://composio.dev/content/pi-vs-omp"}},"_tags":["story","author_hantusk","story_49526765"],"author":"hantusk","created_at":"2026-09-01T19:23:10Z","created_at_i":1788290590,"num_comments":0,"objectID":"49526765","points":1,"story_id":49526765,"title":"Pi vs. OMP (Oh My Pi Agent)","updated_at":"2026-09-01T19:38:37Z","url":"https://composio.dev/content/pi-vs-omp"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"kachapopopow"},"title":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"<em>Oh My PI</em>: coding agent CLI, unified LLM API, TUI and web UI libraries"},"url":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"https://github.com/can1357/<em>oh-my-pi</em>"}},"_tags":["story","author_kachapopopow","story_46664368"],"author":"kachapopopow","created_at":"2026-01-18T02:43:53Z","created_at_i":1768704233,"num_comments":0,"objectID":"46664368","points":1,"story_id":46664368,"title":"Oh My PI: coding agent CLI, unified LLM API, TUI and web UI libraries","updated_at":"2026-03-05T23:27:07Z","url":"https://github.com/can1357/oh-my-pi"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"1nv1n"},"story_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"I wanted to share my side-project: gPTY. Started off as an idea to combine Godot and Rust in a project (two stacks I wanted to use more to learn more). The base inspiration was tmux - simply allow spawning multiple PTYs and then let the user grid/tile them how they see fit.<p>But since we have the Godot game engine at our disposal, we can do some more interesting things, like add an FPS counter, and then subsequently also let people set their preferred FPS (the idea being the potential lower power draw if someone's running it on a laptop on battery power vs someone running it on a desktop with high/native FPS). In its current state, with me using <em>Oh-my-Pi</em> a lot, it's evolving into a terminal workspace that can be used for orchestrating autonomous AI agents by way of dogfooding (or you know, just run herdr inside of gPTY - it's the better orchestrator and just good software - I found it after starting this project, and now I'm finding myself using it a lot).<p>Also, we're not limited to just terminals. Since we have Godot, we have basically a 2D (and potentially a 3D) canvas to play with. We can already full-screen the app for &quot;zen&quot; mode, no taskbar, no distractions. TUI die-hards can have their media or other apps entirely in terminal panes.<p>There has been some ground-work on getting Markdowns displayed properly done and I want to work on some kind of Wiki framework for local knowledge-management next, then create more types of panes (think native audio/video on a media pane, that sits alongside your terminal pane), and some simple 2D games (like snake) to prototype. More details are on the ROADMAP.<p>What's not easy (and probably won't happen) is a browser. Having done a couple of (small) projects using Electron already, the temptation to ditch Godot/Rust (learning curve) did come up (and also the ecosystem, the ease with which I could pull components and use web technologies - development velocity would definitely be higher there). But on the flipside, given all of the available LLM and AI support that we are privileged to have today, I figured the velocity should be comparable depending on how much I leaned on those. And lean I did.<p>Godot/Rust seemed the better call to me and my intent anyway - going with the 'it's not just the end but the journey that matters' philosophy. So yes, there has been heavy use of LLMs &amp; AI to generate a lot of the code. But I do review and steer actively, not relying solely on vibes, and there's a few bits here &amp; there that have been human authored.<p>There are definitely a lot of polish and QoL items that need to land to make the end user experience better, but in the meantime, let me know your thoughts and/or concerns!<p>Repository: <a href=\"https://github.com/godot-pty/gpty\" rel=\"nofollow\">https://github.com/godot-pty/gpty</a><p>Docs/Blog: <a href=\"https://godot-pty.github.io/gpty/\" rel=\"nofollow\">https://godot-pty.github.io/gpty/</a>"},"title":{"matchLevel":"none","matchedWords":[],"value":"Show HN: Godot and Rust based multiplexer (terminal panes and more)"},"url":{"matchLevel":"none","matchedWords":[],"value":"https://github.com/godot-pty/gpty"}},"_tags":["story","author_1nv1n","story_49660676","show_hn"],"author":"1nv1n","children":[49660677,49660948,49661044,49661214,49661263,49661307,49661327,49661705,49661710,49662081,49662113,49662190,49662725,49662845,49662882,49663145,49664340,49665949,49667526,49669020,49669446,49680099],"created_at":"2026-09-11T16:03:05Z","created_at_i":1789142585,"num_comments":50,"objectID":"49660676","points":97,"story_id":49660676,"story_text":"I wanted to share my side-project: gPTY. Started off as an idea to combine Godot and Rust in a project (two stacks I wanted to use more to learn more). The base inspiration was tmux - simply allow spawning multiple PTYs and then let the user grid&#x2F;tile them how they see fit.<p>But since we have the Godot game engine at our disposal, we can do some more interesting things, like add an FPS counter, and then subsequently also let people set their preferred FPS (the idea being the potential lower power draw if someone&#x27;s running it on a laptop on battery power vs someone running it on a desktop with high&#x2F;native FPS). In its current state, with me using Oh-my-Pi a lot, it&#x27;s evolving into a terminal workspace that can be used for orchestrating autonomous AI agents by way of dogfooding (or you know, just run herdr inside of gPTY - it&#x27;s the better orchestrator and just good software - I found it after starting this project, and now I&#x27;m finding myself using it a lot).<p>Also, we&#x27;re not limited to just terminals. Since we have Godot, we have basically a 2D (and potentially a 3D) canvas to play with. We can already full-screen the app for &quot;zen&quot; mode, no taskbar, no distractions. TUI die-hards can have their media or other apps entirely in terminal panes.<p>There has been some ground-work on getting Markdowns displayed properly done and I want to work on some kind of Wiki framework for local knowledge-management next, then create more types of panes (think native audio&#x2F;video on a media pane, that sits alongside your terminal pane), and some simple 2D games (like snake) to prototype. More details are on the ROADMAP.<p>What&#x27;s not easy (and probably won&#x27;t happen) is a browser. Having done a couple of (small) projects using Electron already, the temptation to ditch Godot&#x2F;Rust (learning curve) did come up (and also the ecosystem, the ease with which I could pull components and use web technologies - development velocity would definitely be higher there). But on the flipside, given all of the available LLM and AI support that we are privileged to have today, I figured the velocity should be comparable depending on how much I leaned on those. And lean I did.<p>Godot&#x2F;Rust seemed the better call to me and my intent anyway - going with the &#x27;it&#x27;s not just the end but the journey that matters&#x27; philosophy. So yes, there has been heavy use of LLMs &amp; AI to generate a lot of the code. But I do review and steer actively, not relying solely on vibes, and there&#x27;s a few bits here &amp; there that have been human authored.<p>There are definitely a lot of polish and QoL items that need to land to make the end user experience better, but in the meantime, let me know your thoughts and&#x2F;or concerns!<p>Repository: <a href=\"https:&#x2F;&#x2F;github.com&#x2F;godot-pty&#x2F;gpty\" rel=\"nofollow\">https:&#x2F;&#x2F;github.com&#x2F;godot-pty&#x2F;gpty</a><p>Docs&#x2F;Blog: <a href=\"https:&#x2F;&#x2F;godot-pty.github.io&#x2F;gpty&#x2F;\" rel=\"nofollow\">https:&#x2F;&#x2F;godot-pty.github.io&#x2F;gpty&#x2F;</a>","title":"Show HN: Godot and Rust based multiplexer (terminal panes and more)","updated_at":"2026-09-21T17:05:00Z","url":"https://github.com/godot-pty/gpty"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"jatora"},"story_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"I run a lot of coding agents at once, mostly Claude Code, Codex and pi, across several projects. I got tired of having so many windows open, of agents leaving hung processes behind, of tracking which one needed input, and of having to remote into my desktop or SSH into a single shell to manage them from my phone. Even then, a terminal on a phone couldn't show me the frontend or output I needed to verify the work. I tried Orca and herdr. Both were better than my setup, but neither had a real WYSIWYG markdown notes surface, and neither had a phone interface I could actually work in.<p>swe-mux is the terminal multiplexer I built for that. It runs any shell or CLI in real terminals on your own machine and serves the same live sessions to your desktop and your phone. It's built around agents: Claude Code, Codex, opencode, pi and <em>oh-my-pi</em> get a live status (working, done, or waiting on you), searchable history and resume. Anything else runs as an ordinary terminal with everything except the status.<p>This is the actual terminal and not a UI wrapped around the agent, and you can just use it that way, or start claude in a plain shell and that pane becomes an agent session.<p>On the desktop it's panes and tabs of sessions, notes, files and dev-server previews. Input behaves the same across agents: Shift+Enter, Ctrl+Backspace, paste and click-to-position work in all of them, shortcuts come from presets (tmux, VS Code, Vim, Emacs), and you can paste images straight from the clipboard.<p>Notes and any markdown file in the project open in the same WYSIWYG editor, on desktop and phone, and autosave. No more scattered locations for notes.<p>On the phone it's the same sessions over your own Tailscale network, installed as a PWA: the terminal with a usable keyboard, git diff review, files, notes, and push notifications when an agent needs a decision. A dev server on the host's 127.0.0.1 is proxied through swe-mux, so the phone can open it without exposing another port. Being able to play-test and check a frontend there is what finally closed the development loop on the phone for a lot of my projects. Now I can and do swe anywhere. It's a problem<p>Two things I use constantly: a prompt queue, so something I think of mid-turn waits until the agent is done instead of interrupting it, and messaging between sessions, so a Codex session can ask a Claude session something. I don't want my workflow tied to one provider.<p>General architecture:<p>Sessions survive restarts. A separate supervisor process owns the terminals, so the daemon can restart and the app can update without killing anything. I update it several times a day, so this wasn't optional.<p>Status comes from the CLI's hooks, its transcript and the terminal itself, because no single one reliably tells &quot;waiting on you&quot; apart from &quot;busy in the background&quot;.<p>There's no swe-mux account, no server of mine in the path, and no telemetry. Just your tailnet<p>It was built on Windows, which is where I use it every day and the only platform with a packaged desktop app. The installer is unsigned, so the PyPI install is the way to avoid the SmartScreen prompt. macOS and Linux run the daemon and use a browser. CI runs it on all three, but I haven't lived on those myself, however Linux I have tested.<p>Demo, the real UI with simulated agents, nothing to install: <a href=\"https://swemux.dev/demo/\" rel=\"nofollow\">https://swemux.dev/demo/</a><p>Install: `uv tool install swe-mux` (or `pipx install swe-mux`), then run `swe-mux` on Windows or `swemux start` elsewhere.<p>Apache-2.0: <a href=\"https://github.com/jatoran/swe-mux\" rel=\"nofollow\">https://github.com/jatoran/swe-mux</a><p>I've put a lot of work into the initial user experience to help streamline a quick setup depending on someone's needs, feedback greatly appreciated, thanks"},"title":{"matchLevel":"none","matchedWords":[],"value":"Show HN: swe-mux \u2013 A terminal multiplexer for coding agents optimized for mobile"},"url":{"matchLevel":"none","matchedWords":[],"value":"https://github.com/jatoran/swe-mux"}},"_tags":["story","author_jatora","story_49867905","show_hn"],"author":"jatora","created_at":"2026-09-27T16:02:04Z","created_at_i":1790524924,"num_comments":0,"objectID":"49867905","points":3,"story_id":49867905,"story_text":"I run a lot of coding agents at once, mostly Claude Code, Codex and pi, across several projects. I got tired of having so many windows open, of agents leaving hung processes behind, of tracking which one needed input, and of having to remote into my desktop or SSH into a single shell to manage them from my phone. Even then, a terminal on a phone couldn&#x27;t show me the frontend or output I needed to verify the work. I tried Orca and herdr. Both were better than my setup, but neither had a real WYSIWYG markdown notes surface, and neither had a phone interface I could actually work in.<p>swe-mux is the terminal multiplexer I built for that. It runs any shell or CLI in real terminals on your own machine and serves the same live sessions to your desktop and your phone. It&#x27;s built around agents: Claude Code, Codex, opencode, pi and oh-my-pi get a live status (working, done, or waiting on you), searchable history and resume. Anything else runs as an ordinary terminal with everything except the status.<p>This is the actual terminal and not a UI wrapped around the agent, and you can just use it that way, or start claude in a plain shell and that pane becomes an agent session.<p>On the desktop it&#x27;s panes and tabs of sessions, notes, files and dev-server previews. Input behaves the same across agents: Shift+Enter, Ctrl+Backspace, paste and click-to-position work in all of them, shortcuts come from presets (tmux, VS Code, Vim, Emacs), and you can paste images straight from the clipboard.<p>Notes and any markdown file in the project open in the same WYSIWYG editor, on desktop and phone, and autosave. No more scattered locations for notes.<p>On the phone it&#x27;s the same sessions over your own Tailscale network, installed as a PWA: the terminal with a usable keyboard, git diff review, files, notes, and push notifications when an agent needs a decision. A dev server on the host&#x27;s 127.0.0.1 is proxied through swe-mux, so the phone can open it without exposing another port. Being able to play-test and check a frontend there is what finally closed the development loop on the phone for a lot of my projects. Now I can and do swe anywhere. It&#x27;s a problem<p>Two things I use constantly: a prompt queue, so something I think of mid-turn waits until the agent is done instead of interrupting it, and messaging between sessions, so a Codex session can ask a Claude session something. I don&#x27;t want my workflow tied to one provider.<p>General architecture:<p>Sessions survive restarts. A separate supervisor process owns the terminals, so the daemon can restart and the app can update without killing anything. I update it several times a day, so this wasn&#x27;t optional.<p>Status comes from the CLI&#x27;s hooks, its transcript and the terminal itself, because no single one reliably tells &quot;waiting on you&quot; apart from &quot;busy in the background&quot;.<p>There&#x27;s no swe-mux account, no server of mine in the path, and no telemetry. Just your tailnet<p>It was built on Windows, which is where I use it every day and the only platform with a packaged desktop app. The installer is unsigned, so the PyPI install is the way to avoid the SmartScreen prompt. macOS and Linux run the daemon and use a browser. CI runs it on all three, but I haven&#x27;t lived on those myself, however Linux I have tested.<p>Demo, the real UI with simulated agents, nothing to install: <a href=\"https:&#x2F;&#x2F;swemux.dev&#x2F;demo&#x2F;\" rel=\"nofollow\">https:&#x2F;&#x2F;swemux.dev&#x2F;demo&#x2F;</a><p>Install: `uv tool install swe-mux` (or `pipx install swe-mux`), then run `swe-mux` on Windows or `swemux start` elsewhere.<p>Apache-2.0: <a href=\"https:&#x2F;&#x2F;github.com&#x2F;jatoran&#x2F;swe-mux\" rel=\"nofollow\">https:&#x2F;&#x2F;github.com&#x2F;jatoran&#x2F;swe-mux</a><p>I&#x27;ve put a lot of work into the initial user experience to help streamline a quick setup depending on someone&#x27;s needs, feedback greatly appreciated, thanks","title":"Show HN: swe-mux \u2013 A terminal multiplexer for coding agents optimized for mobile","updated_at":"2026-09-27T23:35:48Z","url":"https://github.com/jatoran/swe-mux"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"opwizardx"},"story_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"I wanted all of my sessions preserved in one S3 storage dump I could access from anywhere. I couldn't find anything that could store sessions from multiple machines in one remote S3 bucket without running a database service anywhere, make them available as an MCP for my agent, and efficiently search through them. pond was born out of this, and out of my necessity to stop being locked to my laptop. I'm convinced that the sessions we generate with agents are one of the most precious assets we have and it still feels so odd to me that we throw them away like we couldn't care less.<p>With pond you just point it at an S3 bucket or a local directory. Under the hood is in-process Lance. It also supports safe concurrent writes, so I just collect all my sessions from all my local machines, remote VMs and telegram agent setups in one bucket dir. Right now that is 14,861 sessions, 2.83M messages, 10.6 GiB, from 8 tools (Claude Code, Codex, opencode, pi, Claude Desktop, <em>oh-my-pi</em> and a few more).<p>The one thing I really want to get optimized is remote storage read speeds. Local queries run in milliseconds to two seconds on my whole corpus, while remote storage is still 20-30 seconds per call. My goal is to get them on par.<p>What I would really want to know is which client to support next, and whether keeping the vector search and all of that embedded-model hassle is worth its keep. FTS-only would have been much faster, and there is hardly any good benchmark I can find that could help me close this question."},"title":{"matchLevel":"none","matchedWords":[],"value":"Show HN: Pond \u2013 lossless archive for agent sessions in your own S3"},"url":{"matchLevel":"none","matchedWords":[],"value":"https://github.com/tenequm/pond"}},"_tags":["story","author_opwizardx","story_49376500","show_hn"],"author":"opwizardx","children":[49391256],"created_at":"2026-08-20T16:04:21Z","created_at_i":1787241861,"num_comments":0,"objectID":"49376500","points":3,"story_id":49376500,"story_text":"I wanted all of my sessions preserved in one S3 storage dump I could access from anywhere. I couldn&#x27;t find anything that could store sessions from multiple machines in one remote S3 bucket without running a database service anywhere, make them available as an MCP for my agent, and efficiently search through them. pond was born out of this, and out of my necessity to stop being locked to my laptop. I&#x27;m convinced that the sessions we generate with agents are one of the most precious assets we have and it still feels so odd to me that we throw them away like we couldn&#x27;t care less.<p>With pond you just point it at an S3 bucket or a local directory. Under the hood is in-process Lance. It also supports safe concurrent writes, so I just collect all my sessions from all my local machines, remote VMs and telegram agent setups in one bucket dir. Right now that is 14,861 sessions, 2.83M messages, 10.6 GiB, from 8 tools (Claude Code, Codex, opencode, pi, Claude Desktop, oh-my-pi and a few more).<p>The one thing I really want to get optimized is remote storage read speeds. Local queries run in milliseconds to two seconds on my whole corpus, while remote storage is still 20-30 seconds per call. My goal is to get them on par.<p>What I would really want to know is which client to support next, and whether keeping the vector search and all of that embedded-model hassle is worth its keep. FTS-only would have been much faster, and there is hardly any good benchmark I can find that could help me close this question.","title":"Show HN: Pond \u2013 lossless archive for agent sessions in your own S3","updated_at":"2026-08-21T17:24:48Z","url":"https://github.com/tenequm/pond"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"mrsirg"},"story_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"yeah everyone is building there own harness. it\u2019s fun though, owning the whole stack. tried opencode, pi, <em>oh my pi</em>, prime agent. wrote some extensions for pi. just wasn\u2019t doing it for me. so I built my own agent runtime. and now it\u2019s turned into my daily driver. Been dogfooding it for a month now and it\u2019s the only harness I want to use. zero bloat, concurrent tool calls, plugins. latest release avail for download (only for macOS/linux).<p>https://mrsirg97-rgb.github.io/rig/\nhttps://github.com/mrsirg97-rgb/rig"},"title":{"matchLevel":"none","matchedWords":[],"value":"Building your own agent harness is fun"}},"_tags":["story","author_mrsirg","story_49667385","ask_hn"],"author":"mrsirg","children":[49667392,49667712],"created_at":"2026-09-12T00:41:30Z","created_at_i":1789173690,"num_comments":2,"objectID":"49667385","points":2,"story_id":49667385,"story_text":"yeah everyone is building there own harness. it\u2019s fun though, owning the whole stack. tried opencode, pi, oh my pi, prime agent. wrote some extensions for pi. just wasn\u2019t doing it for me. so I built my own agent runtime. and now it\u2019s turned into my daily driver. Been dogfooding it for a month now and it\u2019s the only harness I want to use. zero bloat, concurrent tool calls, plugins. latest release avail for download (only for macOS&#x2F;linux).<p>https:&#x2F;&#x2F;mrsirg97-rgb.github.io&#x2F;rig&#x2F;\nhttps:&#x2F;&#x2F;github.com&#x2F;mrsirg97-rgb&#x2F;rig","title":"Building your own agent harness is fun","updated_at":"2026-09-12T02:30:30Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"arush15june"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"I am 4.1 maxxing on commandcode GOAT Plan + api rates with <em>oh my pi</em> for the last 4 weeks, it's absolutely amazing and crazy fast, it's alright if it makes a mistake, I have enough time to iterate again, I have also added an advisor layer of mimo 2.6 pro which does make it a notch smarter. Getting haiku 5.5/sonnet5.5 to work on plans and letting 4.1 flash work through it is helping a ton too.<p>I am a big ChatGPT fan, all our team has ChatGPT Subs, but the TPS across all models including luna is just so damn slow.<p>Commandcode giving 60$ worth of Deepseek for 10$ is just genuinely goat.<p>And it never says no for cyber tasks so that's a big win"},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Why isn't the industry freaking out about DeepSeek 4.1 Flash?"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://www.dgt.is/blog/2026-10-07-deepseek-freek-out/"}},"_tags":["comment","author_arush15june","story_50000488"],"author":"arush15june","comment_text":"I am 4.1 maxxing on commandcode GOAT Plan + api rates with oh my pi for the last 4 weeks, it&#x27;s absolutely amazing and crazy fast, it&#x27;s alright if it makes a mistake, I have enough time to iterate again, I have also added an advisor layer of mimo 2.6 pro which does make it a notch smarter. Getting haiku 5.5&#x2F;sonnet5.5 to work on plans and letting 4.1 flash work through it is helping a ton too.<p>I am a big ChatGPT fan, all our team has ChatGPT Subs, but the TPS across all models including luna is just so damn slow.<p>Commandcode giving 60$ worth of Deepseek for 10$ is just genuinely goat.<p>And it never says no for cyber tasks so that&#x27;s a big win","created_at":"2026-10-08T21:14:02Z","created_at_i":1791494042,"objectID":"50012393","parent_id":50000488,"story_id":50000488,"story_title":"Why isn't the industry freaking out about DeepSeek 4.1 Flash?","story_url":"https://www.dgt.is/blog/2026-10-07-deepseek-freek-out/","updated_at":"2026-10-08T21:30:56Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"jeremyjh"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"I use <em>oh-my-pi</em> (omp.sh) - mostly with stock settings and skills but its a &quot;fully loaded&quot; harness that you don't really have to add anything to. I do change a couple of things: I set it to prefer subagents, and I enable rewind. It is crazy that rewind is not a default, it saves a TON of context - if the agent goes down a crazy path that burns a lot of tokens it can rewind to an earlier checkpoint with the exploration or bug hunting summary.<p>Presently I'm using sol-high for the default agent which does orchestration and a lot of smaller investigation and coding tasks itself.  sol-max for planning and review. Luna-max for planned coding and general tasks. I also have a $10 minimax plan and use M3 for exploration and library roles, but I could probably be using Luna for that just as well and still only very rarely run into usage issues.<p>I don't use any plugins or skill libraries apart from Caveman and I'm not sure how useful that really is anymore so I'd start without it so you have a baseline to compare. I do think it reduces context usage a bit but I haven't measured it recently. Caveman also includes some team, agent &amp; investigation skills - again they might be helping but I haven't re-evaluated since like 90 days ago."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Claude Haiku 5.5"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://www.anthropic.com/claude-haiku-5-5"}},"_tags":["comment","author_jeremyjh","story_49996437"],"author":"jeremyjh","comment_text":"I use oh-my-pi (omp.sh) - mostly with stock settings and skills but its a &quot;fully loaded&quot; harness that you don&#x27;t really have to add anything to. I do change a couple of things: I set it to prefer subagents, and I enable rewind. It is crazy that rewind is not a default, it saves a TON of context - if the agent goes down a crazy path that burns a lot of tokens it can rewind to an earlier checkpoint with the exploration or bug hunting summary.<p>Presently I&#x27;m using sol-high for the default agent which does orchestration and a lot of smaller investigation and coding tasks itself.  sol-max for planning and review. Luna-max for planned coding and general tasks. I also have a $10 minimax plan and use M3 for exploration and library roles, but I could probably be using Luna for that just as well and still only very rarely run into usage issues.<p>I don&#x27;t use any plugins or skill libraries apart from Caveman and I&#x27;m not sure how useful that really is anymore so I&#x27;d start without it so you have a baseline to compare. I do think it reduces context usage a bit but I haven&#x27;t measured it recently. Caveman also includes some team, agent &amp; investigation skills - again they might be helping but I haven&#x27;t re-evaluated since like 90 days ago.","created_at":"2026-10-08T11:18:22Z","created_at_i":1791458302,"objectID":"50004447","parent_id":50001004,"story_id":49996437,"story_title":"Claude Haiku 5.5","story_url":"https://www.anthropic.com/claude-haiku-5-5","updated_at":"2026-10-08T11:23:25Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"haellsigh"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"It looks like he uses <em>oh-my-pi</em>, which is a full featured version of pi and has this to-do style out of the box."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://github.com/Niko1221/Strata"}},"_tags":["comment","author_haellsigh","story_49953495"],"author":"haellsigh","comment_text":"It looks like he uses oh-my-pi, which is a full featured version of pi and has this to-do style out of the box.","created_at":"2026-10-05T07:01:48Z","created_at_i":1791183708,"objectID":"49961552","parent_id":49961324,"story_id":49953495,"story_title":"Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s","story_url":"https://github.com/Niko1221/Strata","updated_at":"2026-10-05T07:03:27Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"Winfred-zz"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"I just ran a set of benchmarks, ninfer-3090-qwen3.8-27b (so mix of Q4 and Q5) vs strata-qwen3.8-flash-next-iq3_xxs (so Q3):<p>\u2502-------------------          \u2502 Ninfer-3090    \u2502 Strata<p>\u2502 Code generation          \u2502 52/78 (66.7%)  \u2502 70/78 (89.7%)<p>\u2502 Code completion          \u2502 40/50 (80.0%)  \u2502 44/50 (88.0%)<p>\u2502 Total-------------                    \u2502 92/128 (71.9%) \u2502 114/128 (89.1%)<p>\u2502 API failures------ \u2502 10             \u2502 5<p>- Ninfer generation: ~122 min total.<p>- Strata generation: ~142 min total.<p>So strata is a little slower, but keep in mind that ninfer-3090 is very optimized for a Qwen 3.8. Standard Qwen 3.8 runs at 20 t/s, this modified version can do 50 t/s (but it's <i>extremely long in it's thinking, it just goes on and on</i>.<p>This is on a 3090 that will crash unless power capped, with a Zen 2 CPU, 64GB DDR4 with a PCIe that refuses to go higher than 8x (basically pretty crappy all in all).<p>Yet with some tweaking and optimizing I still manage to get strata to run at 40 to 60 t/s.<p>That strata has been optimized on my <em>Oh My Pi</em> conversations. So when I'm using it, it's probably faster and closer to ninfer in speed than during those unoptimized benchmark tests."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://github.com/Niko1221/Strata"}},"_tags":["comment","author_Winfred-zz","story_49953495"],"author":"Winfred-zz","children":[49958674,49959827,49960046,49961097,49961154],"comment_text":"I just ran a set of benchmarks, ninfer-3090-qwen3.8-27b (so mix of Q4 and Q5) vs strata-qwen3.8-flash-next-iq3_xxs (so Q3):<p>\u2502-------------------          \u2502 Ninfer-3090    \u2502 Strata<p>\u2502 Code generation          \u2502 52&#x2F;78 (66.7%)  \u2502 70&#x2F;78 (89.7%)<p>\u2502 Code completion          \u2502 40&#x2F;50 (80.0%)  \u2502 44&#x2F;50 (88.0%)<p>\u2502 Total-------------                    \u2502 92&#x2F;128 (71.9%) \u2502 114&#x2F;128 (89.1%)<p>\u2502 API failures------ \u2502 10             \u2502 5<p>- Ninfer generation: ~122 min total.<p>- Strata generation: ~142 min total.<p>So strata is a little slower, but keep in mind that ninfer-3090 is very optimized for a Qwen 3.8. Standard Qwen 3.8 runs at 20 t&#x2F;s, this modified version can do 50 t&#x2F;s (but it&#x27;s <i>extremely long in it&#x27;s thinking, it just goes on and on</i>.<p>This is on a 3090 that will crash unless power capped, with a Zen 2 CPU, 64GB DDR4 with a PCIe that refuses to go higher than 8x (basically pretty crappy all in all).<p>Yet with some tweaking and optimizing I still manage to get strata to run at 40 to 60 t&#x2F;s.<p>That strata has been optimized on my Oh My Pi conversations. So when I&#x27;m using it, it&#x27;s probably faster and closer to ninfer in speed than during those unoptimized benchmark tests.","created_at":"2026-10-04T21:34:00Z","created_at_i":1791149640,"objectID":"49958120","parent_id":49955565,"story_id":49953495,"story_title":"Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s","story_url":"https://github.com/Niko1221/Strata","updated_at":"2026-10-07T02:26:32Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"zahrevsky"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"There is a command in <em>oh-my-pi</em> called &quot;/omfg &lt;problem&gt;&quot;. You explain what is wrong with agent's response, and it writes a hook to make sure that the problem doesn't happen again. It then re-runs your previous prompt to make sure that hook is triggered, and if not, it rewrites the hook to make your previous prompt trigger the hook. Then each next agent's response is checked by the hook, and if it is triggered, the agent receives feedback on what's wrong and what must be done differently."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Agents don't need memory, they need documentation"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://liao.gg/blog/agents-dont-need-memory"}},"_tags":["comment","author_zahrevsky","story_49945933"],"author":"zahrevsky","children":[49950577,49951134,49952632],"comment_text":"There is a command in oh-my-pi called &quot;&#x2F;omfg &lt;problem&gt;&quot;. You explain what is wrong with agent&#x27;s response, and it writes a hook to make sure that the problem doesn&#x27;t happen again. It then re-runs your previous prompt to make sure that hook is triggered, and if not, it rewrites the hook to make your previous prompt trigger the hook. Then each next agent&#x27;s response is checked by the hook, and if it is triggered, the agent receives feedback on what&#x27;s wrong and what must be done differently.","created_at":"2026-10-04T00:49:46Z","created_at_i":1791074986,"objectID":"49949426","parent_id":49949330,"story_id":49945933,"story_title":"Agents don't need memory, they need documentation","story_url":"https://liao.gg/blog/agents-dont-need-memory","updated_at":"2026-10-08T12:33:38Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"ttoinou"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"Ive been using this since it was initially released with deepseek v4 flash, and it is absolutely the best launcher ever on my m5 max 128gb<p>Now Ive been running qwen 3.8 flash next for more than a week and it\u2019s doing great, really fast and super long context windows. Sometimes the model is behaving stupidly by not remembering something I said earlier but it could be also a problem from the agentic AI harness. Im using <em>oh my pi</em> but Im wondering what people are using ds4 with here ?"},"story_title":{"matchLevel":"none","matchedWords":[],"value":"From the creator of Redis; run LLM locally with ds4"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://dwarfstar.sh/"}},"_tags":["comment","author_ttoinou","story_49936575"],"author":"ttoinou","children":[49941769],"comment_text":"Ive been using this since it was initially released with deepseek v4 flash, and it is absolutely the best launcher ever on my m5 max 128gb<p>Now Ive been running qwen 3.8 flash next for more than a week and it\u2019s doing great, really fast and super long context windows. Sometimes the model is behaving stupidly by not remembering something I said earlier but it could be also a problem from the agentic AI harness. Im using oh my pi but Im wondering what people are using ds4 with here ?","created_at":"2026-10-02T22:19:03Z","created_at_i":1790979543,"objectID":"49939328","parent_id":49936575,"story_id":49936575,"story_title":"From the creator of Redis; run LLM locally with ds4","story_url":"https://dwarfstar.sh/","updated_at":"2026-10-05T07:15:12Z"},{"_highlightResult":{"author":{"matchLevel":"none","matchedWords":[],"value":"dom96"},"comment_text":{"fullyHighlighted":false,"matchLevel":"full","matchedWords":["oh","my","pi"],"value":"I\u2019ve tried Pi and honestly it felt pretty rough and raw. Kind of like Arch Linux. Even with <em>oh-my-pi</em> the experience just didn\u2019t seem polished enough for my taste. I\u2019m not really sure why I would use it over OpenCode."},"story_title":{"matchLevel":"none","matchedWords":[],"value":"Pi 1.0"},"story_url":{"matchLevel":"none","matchedWords":[],"value":"https://earendil.com/posts/pi-1-0/"}},"_tags":["comment","author_dom96","story_49926069"],"author":"dom96","children":[49932745],"comment_text":"I\u2019ve tried Pi and honestly it felt pretty rough and raw. Kind of like Arch Linux. Even with oh-my-pi the experience just didn\u2019t seem polished enough for my taste. I\u2019m not really sure why I would use it over OpenCode.","created_at":"2026-10-02T12:17:35Z","created_at_i":1790943455,"objectID":"49932649","parent_id":49926069,"story_id":49926069,"story_title":"Pi 1.0","story_url":"https://earendil.com/posts/pi-1-0/","updated_at":"2026-10-02T15:33:18Z"}],"hitsPerPage":20,"nbHits":236,"nbPages":12,"page":0,"params":"query=oh-my-pi&advancedSyntax=true&analyticsTags=backend","processingTimeMS":26,"processingTimingsMS":{"_request":{"roundTrip":23},"afterFetch":{"format":{"highlighting":1,"total":1}},"fetch":{"scanning":24,"total":25},"total":26},"query":"oh-my-pi","serverTimeMS":28}
