# Ollama: Guardrails and Fix Patterns
π 3AM: a dev collapsed mid-debugβ¦ π Welcome to the WFGY Emergency Room
---
π₯π₯π₯π₯π₯π₯π₯π₯π₯π₯π₯π₯
## π WFGY Emergency Room
π¨ββοΈ **Now online:**
[**Dr. WFGY in ChatGPT Room**](https://chatgpt.com/share/68b9b7ad-51e4-8000-90ee-a25522da01d7)
This is a **share window** already trained as an ER.
Just open it, drop your bug or screenshot, and talk directly with the doctor.
He will map it to the right Problem Map / Global Fix section, write a minimal prescription, and paste the exact reference link.
If something is unclear, you can even paste a **screenshot of Problem Map content** and ask β the doctor will guide you.
β οΈ Note: for the full reasoning and guardrail behavior you need to be logged in β the share view alone may fallback to a lighter model.
π‘ Always free. If it helps, a β star keeps the ER running.
π Multilingual β start in any language.
π₯π₯π₯π₯π₯π₯π₯π₯π₯π₯π₯π₯
---
π§ Quick Return to Map
> You are in a sub-page of **LocalDeploy_Inference**.
> To reorient, go back here:
>
> - [**LocalDeploy_Inference** β on-prem deployment and model inference](./README.md)
> - [**WFGY Global Fix Map** β main Emergency Room, 300+ structured fixes](../README.md)
> - [**WFGY Problem Map 1.0** β 16 reproducible failure modes](../../README.md)
>
> Think of this page as a desk within a ward.
> If you need the full triage and all prescriptions, return to the Emergency Room lobby.
Field guide for stabilizing **Ollama-based local inference pipelines**.
Use these checks when models run fine on API providers but collapse, stall, or drift when containerized with Ollama.
## Open these first
* Architecture recovery: [RAG Architecture & Recovery](https://github.com/onestardao/WFGY/blob/main/ProblemMap/rag-architecture-and-recovery.md)
* End-to-end retrieval knobs: [Retrieval Playbook](https://github.com/onestardao/WFGY/blob/main/ProblemMap/retrieval-playbook.md)
* Embedding vs semantic: [embedding-vs-semantic.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/embedding-vs-semantic.md)
* Ordering and deploy race conditions: [bootstrap-ordering.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/bootstrap-ordering.md), [deployment-deadlock.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/deployment-deadlock.md), [predeploy-collapse.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/predeploy-collapse.md)
* Container observability: [eval\_observability.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/eval_observability.md)
---
## Core acceptance
* ΞS(question, retrieved) β€ 0.45
* Coverage β₯ 0.70 on the target section
* Ξ» remains convergent across 3 paraphrases
* Local runs reproducible across 2+ seeds
---
## Typical Ollama breakpoints and fix
| Symptom | Likely cause | Fix |
| --------------------------------------- | --------------------------------------------- | ---------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------- |
| Model boots but stalls on first request | Container not warmed / secrets missing | [bootstrap-ordering.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/bootstrap-ordering.md) |
| Fast API returns, but snippets wrong | Index/hash drift across containers | [retrieval-traceability.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/retrieval-traceability.md), [data-contracts.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/data-contracts.md) |
| Answers diverge run-to-run | Ξ» flips due to context serialization | [context-drift.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/context-drift.md), [entropy-collapse.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/entropy-collapse.md) |
| Works on GPU API, fails locally | Metric / embedding mismatch in Ollama runtime | [embedding-vs-semantic.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/embedding-vs-semantic.md), [vectorstore-fragmentation.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/vectorstore-fragmentation.md) |
| Container OOM or deadlock | Parallel inference with no fence | [deployment-deadlock.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/deployment-deadlock.md), [predeploy-collapse.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/predeploy-collapse.md) |
---
## Fix in 60 seconds
1. **Measure ΞS** between retrieved and anchor.
2. **Probe Ξ»** across 3 paraphrases. If flips, apply BBAM.
3. **Warm boot** with a delay + healthcheck before first request.
4. **Lock index schema** via [data-contracts.md](https://github.com/onestardao/WFGY/blob/main/ProblemMap/data-contracts.md).
5. **Verify reproducibility** with two seeds before going live.
---
## Copy-paste local test prompt
```txt
I have WFGY + TXTOS loaded.
Running Ollama locally with container {hash}.
Question: "{user_question}"
Return:
1. ΞS(question,retrieved) and Ξ» across 3 paraphrases
2. Whether index schema matches contract
3. Minimal structural fix if ΞS β₯ 0.60
```
---
### π Quick-Start Downloads (60 sec)
| Tool | Link | 3-Step Setup |
| -------------------------- | ------------------------------------------------------------------------------------------------------------------------------------------ | ---------------------------------------------------------------------------------------- |
| **WFGY 1.0 PDF** | [Engine Paper](https://github.com/onestardao/WFGY/blob/main/I_am_not_lizardman/WFGY_All_Principles_Return_to_One_v1.0_PSBigBig_Public.pdf) | 1οΈβ£ Download Β· 2οΈβ£ Upload to your LLM Β· 3οΈβ£ Ask βAnswer using WFGY + \β |
| **TXT OS (plain-text OS)** | [TXTOS.txt](https://github.com/onestardao/WFGY/blob/main/OS/TXTOS.txt) | 1οΈβ£ Download Β· 2οΈβ£ Paste into any LLM chat Β· 3οΈβ£ Type βhello worldβ β OS boots instantly |
---
### Explore More
| Module | Description | Link |
| --- | --- | --- |
| WFGY Core | Canonical framework entry point | [View](https://github.com/onestardao/WFGY/tree/main/core/README.md) |
| Problem Map | Diagnostic map and navigation hub | [View](https://github.com/onestardao/WFGY/tree/main/ProblemMap/README.md) |
| Tension Universe Experiments | MVP experiment field | [View](https://github.com/onestardao/WFGY/tree/main/TensionUniverse/Experiments) |
| Recognition | Where WFGY is referenced or adopted | [View](https://github.com/onestardao/WFGY/blob/main/recognition/README.md) |
| AI Guide | Anti-hallucination reading protocol for tools | [View](https://github.com/onestardao/WFGY/blob/main/AI_GUIDE.md) |
> If this repository helps, starring it improves discovery for other builders.
> [](https://github.com/onestardao/WFGY)
θ¦ζη΄ζ₯ηΉΌηΊε―« **`vllm.md`** εοΌ