Commit Graph

  • 993acc7504 model: update lfm2 parser/renderer for optional thinking (#16359) Jeffrey Morgan 2026-06-14 20:37:08 -07:00
  • 7ea692cb2b llama: update llama.cpp to b9637 (#16609) Jeffrey Morgan 2026-06-14 20:05:08 -07:00
  • e455ea5c66 Reject prompts over the available context length jmorganca/context-shift-default-on Jeffrey Morgan 2026-06-13 22:39:25 -07:00
  • b8f1ed0240 Simplify context shift resolution Jeffrey Morgan 2026-06-13 19:24:04 -07:00
  • 33df53889d Order llama.cpp compat patches jmorganca/update-llamacpp Jeffrey Morgan 2026-06-13 12:06:34 -07:00
  • a20476c722 Revert "Bundle llama.cpp UI assets" Jeffrey Morgan 2026-06-13 12:00:18 -07:00
  • a2ca874de5 Bundle llama.cpp UI assets Jeffrey Morgan 2026-06-13 11:57:50 -07:00
  • 6efd7a6ece Pin llama.cpp to b9626 Jeffrey Morgan 2026-06-13 11:44:29 -07:00
  • 4f5cc32150 llm: fix context limit conditions jmorganca/context-limit-fixes Jeffrey Morgan 2026-06-12 18:03:05 -07:00
  • 821a45ca6e agent: update compaction and context tracking ParthSareen 2026-06-12 14:28:13 -07:00
  • 676facf043 agent: update chat tui controls ParthSareen 2026-06-12 14:13:10 -07:00
  • 50184444da agent: add chat tui controls and legacy aliases ParthSareen 2026-06-12 11:57:44 -07:00
  • 632bd88937 agent: improve chat input controls ParthSareen 2026-06-12 10:42:35 -07:00
  • 2ba0f9b96f agent: add output copy commands ParthSareen 2026-06-12 10:31:46 -07:00
  • 5625904153 agent: window slash command suggestions ParthSareen 2026-06-12 08:12:04 -07:00
  • 96f82b5ff2 agent: add model picker slash command ParthSareen 2026-06-12 00:48:59 -07:00
  • 9da4d0c6f8 agent: restore web tool auto approval and TUI scrolling ParthSareen 2026-06-12 00:33:34 -07:00
  • edd30d7194 docs: add agent improvements report ParthSareen 2026-06-11 21:03:47 -07:00
  • f417732279 agent: bound markdown and compaction caches ParthSareen 2026-06-11 21:02:40 -07:00
  • 72d18c99af agent: validate bash working directory ParthSareen 2026-06-11 21:00:21 -07:00
  • 02befa963c agent: require approval for web tools ParthSareen 2026-06-11 20:58:31 -07:00
  • a743721ec8 agent: harden bash approval classification ParthSareen 2026-06-11 20:56:18 -07:00
  • 2b54b72207 tui: split chat implementation ParthSareen 2026-06-11 20:54:22 -07:00
  • 30f206e235 agent: merge model-aware chat store ParthSareen 2026-06-11 20:45:35 -07:00
  • 13006826f8 agent: batch streaming persistence ParthSareen 2026-06-11 20:42:38 -07:00
  • 76eb5a9dc9 tui: track live working directory ParthSareen 2026-06-11 20:39:35 -07:00
  • cb8af634b3 tui: surface approvals over modals ParthSareen 2026-06-11 20:36:57 -07:00
  • 66cdf46e7d cmd: remove orphaned repl ParthSareen 2026-06-11 20:35:13 -07:00
  • a7c8484bea agent: persist chat images ParthSareen 2026-06-11 20:29:48 -07:00
  • 3be6794929 agent: unify approval dispatch ParthSareen 2026-06-11 19:07:27 -07:00
  • 444af2a712 feat: promote agent loop to canonical run path ParthSareen 2026-06-11 18:48:59 -07:00
  • 6e88139529 models: add cohere2_moe (Command A / North) to the MLX engine jmorganca/cohere-moe Jeffrey Morgan 2026-06-10 11:18:11 -07:00
  • 12e04379cd launch: Fix launch provider drift (#16683) Parth Sareen 2026-06-11 17:21:46 -07:00
  • f8a48df24d llm: decouple prompt caching from context shift (#16639) Parafee41 2026-06-12 07:05:24 +08:00
  • 82e0ddb6fe mlxrunner: harden linear/embedding layers against over-promotion (#16682) Patrick Devine 2026-06-11 13:56:25 -07:00
  • ef4e66bd5f rm extra infromation hoyyeva/qwen-doc Parth Sareen 2026-06-03 14:34:54 -07:00
  • 8bae6e2569 docs: add Qwen Code integration Eva Ho 2026-06-01 15:34:45 -07:00
  • 1abd56b6e6 mlxrunner: record committed MTP drafts before streaming them Jesse Gross 2026-06-01 11:00:46 -07:00
  • ded2db7d86 mlxrunner: capture prefill snapshots across the forward Jesse Gross 2026-05-29 12:12:10 -07:00
  • d00622060f mlxrunner: drive MTP speculation through cache snapshots Jesse Gross 2026-05-29 12:11:30 -07:00
  • 177aefb8a9 nn/recurrent: return per-boundary states from the gated-delta kernels Jesse Gross 2026-05-29 12:07:46 -07:00
  • 07588c64ee mlxrunner/cache: split KVCache and RotatingKVCache into their own files Jesse Gross 2026-05-29 13:16:27 -07:00
  • 4c97a940ca mlxthread: preserve the original stack when worker work panics Jesse Gross 2026-05-27 13:07:33 -07:00
  • 9a251bd11b Update VS Code integration docs for Ollama extension hoyyeva/vscode-docs Eva Ho 2026-06-08 22:54:49 -04:00
  • 74cbf1d2c2 docs: omp (#16552) Bruce MacDonald 2026-06-08 11:43:51 -07:00
  • 5c1e37eb67 docs: hermes desktop (#16549) Bruce MacDonald 2026-06-08 11:43:11 -07:00
  • b990571f17 Fix Windows MLX dl.dll install jmorganca/remove-mlx-imagegen-code Jeffrey Morgan 2026-06-07 14:53:07 -07:00
  • c8651c9692 Clarify unsupported image generation error Jeffrey Morgan 2026-06-07 14:02:21 -07:00
  • 113e57a7cf Clarify unsupported image generation error Jeffrey Morgan 2026-06-07 14:00:24 -07:00
  • 66fc68a5b1 Fix MLX CMake version path Jeffrey Morgan 2026-06-07 13:53:42 -07:00
  • b613f26abb Simplify model management code Jeffrey Morgan 2026-06-07 13:39:10 -07:00
  • f0078ae476 docs: update docs examples to use Gemma 4 instead of Gemma 3 (#16607) Jeffrey Morgan 2026-06-07 12:43:13 -07:00
  • 09a574da83 Pin llama.cpp to b9550 Jeffrey Morgan 2026-06-07 11:51:51 -07:00
  • 2d9b3ffcaa Update llama.cpp to f0156d14 Jeffrey Morgan 2026-06-07 11:26:18 -07:00
  • 96201a623a Add AGENTS.md and CLAUDE.md to root repository (#16604) Jeffrey Morgan 2026-06-07 10:57:59 -07:00
  • 9c94c2b11e docs: describe llama.cpp update process (#16603) Daniel Hiltgen 2026-06-07 10:27:47 -07:00
  • e09b3f9fb5 openai: align models list with tags (#16556) Parth Sareen 2026-06-05 17:59:05 -07:00
  • a0099da2d1 launch: use native Windows Hermes config path (#16558) Bruce MacDonald 2026-06-05 17:29:19 -07:00
  • 25e0e81e12 docs: update Zod example to use native toJSONSchema (#14746) Chris Chen 2026-06-06 09:21:07 +10:00
  • d107e87320 docs: omp brucemacd/omp-doc Bruce MacDonald 2026-06-05 11:48:48 -07:00
  • 2c293172ca docs: hermes desktop brucemacd/hermes-desktop-doc Bruce MacDonald 2026-06-05 11:25:57 -07:00
  • 87cff95af8 launch: oh-my-pi (#16410) Bruce MacDonald 2026-06-04 17:49:49 -07:00
  • 3ef69ef784 mlx: allow the embedding layer to use the nvfp4 global scale (#16527) Patrick Devine 2026-06-04 17:40:01 -07:00
  • 1a7786be14 docs: add cloud model retirement (#16528) Michael Yang 2026-06-04 15:18:38 -07:00
  • 3370ff8b1c launch: hermes-desktop app (#16516) Bruce MacDonald 2026-06-04 11:51:36 -07:00
  • 455f57457d llama.cpp version update (#16511) Daniel Hiltgen 2026-06-04 08:20:57 -07:00
  • 1d955ed990 integrations: hermes windows install (#16487) Bruce MacDonald 2026-06-03 17:40:45 -07:00
  • d071237131 docs: add Cline CLI integration doc (#16341) Eva H 2026-06-03 20:30:01 -04:00
  • 54e32d34e7 launch: hermes-desktop app brucemacd/hermes-desktop-launch Bruce MacDonald 2026-06-03 17:08:52 -07:00
  • 229a1303fb llama-server: fix gemma4 patch wiring (#16477) Daniel Hiltgen 2026-06-03 14:41:03 -07:00
  • ac3d0657a2 launch: migrate pi (#16213) Parth Sareen 2026-06-03 14:35:32 -07:00
  • f634ea374a launch: set copilot token length defaults parth-copilot-token-length-defaults ParthSareen 2026-06-03 14:27:30 -07:00
  • 01557ff313 llama-server: allow GPU offload for projectors (#16473) Daniel Hiltgen 2026-06-03 13:58:40 -07:00
  • 39fc1e1816 launch/opencode: re-apply on enriched model inventory hoyyeva/opencode-thinking Eva Ho 2026-06-03 13:23:35 -07:00
  • 05c684f49f test: dedupe buildModelEntries coverage Eva Ho 2026-05-12 16:09:36 -04:00
  • 9f850a4285 launch/opencode: keep gpt-oss alias toggle behavior Eva Ho 2026-05-11 16:33:03 -04:00
  • 1637945a02 launch/opencode: bound thinking show probes Eva Ho 2026-05-11 16:15:37 -04:00
  • 95014117f0 add test Eva Ho 2026-04-08 15:09:17 -07:00
  • e4ee912421 launch: add thinking capability detection to opencode Eva Ho 2026-04-08 14:42:57 -07:00
  • e5a38739b4 mlx: "requires" in modelfile is being ignored for mlx based models (#16469) Patrick Devine 2026-06-03 13:10:57 -07:00
  • 5f56a289b3 server: classify mmproj GGUFs as projector layers (#16472) Jeffrey Morgan 2026-06-03 12:59:34 -07:00
  • ad8cda255d launch: clean legacy codex profile before launch (#16467) Eva H 2026-06-03 14:49:31 -04:00
  • 1ec8654383 launch: add pool support on windows hoyyeva/poolside-windows Eva Ho 2026-06-03 11:35:59 -07:00
  • 3e1b4fe39d Kill llama-server during Windows cleanup (#16458) Daniel Hiltgen 2026-06-03 10:25:12 -07:00
  • 52196f1a97 llama.cpp version update (#16463) Daniel Hiltgen 2026-06-03 10:20:30 -07:00
  • 50bbda5660 models: add support for gemma4-12b (#16457) Patrick Devine 2026-06-03 07:44:57 -07:00
  • 4b5bdd3b25 fix laguna patch build breakage (#16445) Daniel Hiltgen 2026-06-02 16:35:19 -07:00
  • e828061b6e llm: ignore llama-server SSE ping comments (#16443) Daniel Hiltgen 2026-06-02 15:40:14 -07:00
  • 7a2073d17b docs: configure hermes desktop app (#16440) Bruce MacDonald 2026-06-02 14:32:10 -07:00
  • c952708169 llama: add laguna (poolside) arch via a llama.cpp patch under llama/c… (#16396) Daniel Hiltgen 2026-06-02 13:17:08 -07:00
  • f57d111754 launch: isolate Codex launch configuration (#16437) Parth Sareen 2026-06-02 12:10:46 -07:00
  • c34a79a373 llama.cpp version update (#16426) Daniel Hiltgen 2026-06-02 11:46:56 -07:00
  • b051c9cf83 More harden app markdown URL handling (#16436) Daniel Hiltgen 2026-06-02 11:46:14 -07:00
  • b7b7fa0454 llm: detect llama-server load stalls from output (#16427) Daniel Hiltgen 2026-06-02 11:30:48 -07:00
  • 4c076813be discover: allow Radeon 8060S iGPU by default (#16429) Daniel Hiltgen 2026-06-02 11:15:01 -07:00
  • 6780f0416a Harden app markdown URL handling (#16380) Daniel Hiltgen 2026-06-02 11:14:36 -07:00
  • 35fa277fa9 llm: include cached prompt tokens in llama-server counts (#16428) Daniel Hiltgen 2026-06-02 10:51:01 -07:00
  • 05747b02ab launch: fix opencode local model limits (#16425) Daniel Hiltgen 2026-06-02 10:50:35 -07:00
  • 6dea455b2d remove docs for now brucemacd/omp Bruce MacDonald 2026-06-01 17:28:09 -07:00
  • d44fdea1ef integrations: oh-my-pi Bruce MacDonald 2026-06-01 12:58:54 -07:00