| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Download Repo ZIP] [Original HTTPS Page] |
| Name | Name | Last commit date | ||
|---|---|---|---|---|
parent directory.. | ||||
These examples demonstrate Zerfoo's core value: embeddable ML inference in pure Go. Each example is a standalone program you can build and run with go build.
zerfoo pull google/gemma-3-1b-it-qat-q4_0-ggufOr download directly:
# The model file will be cached in ~/.cache/zerfoo/
zerfoo pull gemma-3-1b-q4| Example | Description | Prerequisites |
|---|---|---|
| chat/ | Interactive chatbot CLI. Demonstrates the zerfoo.Load and model.Chat one-line API with a readline loop. | GGUF model file |
| embedding-search/ | Semantic search over a document corpus using model embeddings and cosine similarity. | GGUF model file |
| rag/ | Retrieval-augmented generation: embed documents, retrieve relevant ones, and generate grounded answers. | GGUF model file |
| code-completion/ | Generate code completions from partial code snippets using inference.LoadFile and model.Generate. | GGUF model file |
| summarization/ | Summarize text from a string or file using a language model. | GGUF model file |
| translation/ | Translate text between languages using a multilingual model. | GGUF model file |
| classification/ | Text classification with grammar-constrained JSON output using inference.WithGrammar. | GGUF model file |
| vision-analysis/ | Analyze images using a vision-capable model with inference.Message.Images. | Vision GGUF model + image |
| audio-transcription/ | Speech-to-text using the OpenAI-compatible /v1/audio/transcriptions endpoint. | Whisper GGUF model + audio file |
| agentic-tool-use/ | Function calling (tool use) with zerfoo.WithTools for agentic AI patterns. | GGUF model file |
| Example | Description | Prerequisites |
|---|---|---|
| inference/ | Load a GGUF model and generate text with sampling options and token streaming. | GGUF model file |
| streaming/ | Streaming chat generation using model.ChatStream with per-token output. | GGUF model file |
| embedding/ | Embed inference inside a custom Go HTTP handler for concurrent request serving. | GGUF model file |
| api-server/ | Start an OpenAI-compatible HTTP server with serve.NewServer and graceful shutdown. | GGUF model file |
| json-output/ | Grammar-guided decoding that constrains output to valid JSON matching a schema. | GGUF model file |
| fine-tuning/ | LoRA fine-tuning of a tabular model: pre-train, adapt, merge, save/load. | None (synthetic data) |
# Build and run the chat example
go build -o chat ./examples/chat/
./chat --model path/to/model.gguf
# Build and run the code completion example
go build -o code-completion ./examples/code-completion/
./code-completion --model path/to/model.gguf --code "func fibonacci(n int) int {"
# With GPU acceleration
./code-completion --model path/to/model.gguf --device cuda --code "func add(a, b int) int {"See docs/getting-started.md for a full tutorial covering CLI usage, library API, and the OpenAI-compatible server.
| Back | FazBrowse Home | New Git URL |