feat: make local GGUF setup resumable and restart-safe - #98
Merged
Merged
Conversation
XYAIStudio
force-pushed
the
codex/local-runtime-sidecar
branch
from
September 28, 2026 16:27
6371812 to
2dd4686
Compare
XYAIStudio
changed the base branch from
codex/local-model-auto-routing-plan
to
develop
September 29, 2026 07:57
Merged
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
FreeOS users can now start from a computer with neither Ollama nor an existing GGUF file. Settings → Models presents a pinned starter catalog, downloads the selected official Qwen GGUF with live progress, verifies its SHA-256, starts the bundled llama.cpp runtime, benchmarks it, and selects it for local chat only after a successful test.
This update closes the reliability gaps needed for the 0.0.7 flow:
.partfiles after cancellation, network failure, or application restart.The catalog currently offers pinned official Apache-2.0 Qwen2.5 1.5B and 3B Q4_K_M builds. The runtime remains pinned and checksum-verified, binds only to
127.0.0.1:11435, uses an argument vector rather than a shell command, logs under the FreeOS home directory, and stops only the process owned by FreeOS. Ollama remains an optional runtime and import path.Validation:
PYTHONUTF8=1 uv run pytest -n 4 -m "not live"— 3588 passed, 114 skippeduv run mypy src/octop— no issues in 590 source filesuv run ruff check src testsanduv run ruff format --check src tests— 1158 files formattedtsc --noEmitand production build passedThis PR remains stacked on #97 so the implementation and architecture plan can be reviewed separately. #97 now targets
developunder the repository release policy.