Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 20:13:24 EDT

Explore

PostsPeople
LatestRanked
Load more
@yzma.aiOct 10, 2026, 8:52 AM

Jeyzma is like @typesafeai.bsky.social minus $870 million plus @golang.org & @tinygo.org & @hf.co llama.cpp

Go make a decision!

Already available right now in your browser.

jeyzma.com

#jev #decisionAI #browserAI #llamacpp #systemOne

@dark-city-beta.bsky.socialOct 9, 2026, 3:49 PM

🧠 Second Hemisphere for Local AI Agents: Beating Context Overflow

During long-horizon runs (coding, bash, 50–100+ steps), tool outputs saturate the context window → HTTP 400 → session dies.

🔗 github.com/Dark-city-be...

#OpenSource #AIagents #LocalAI #llamaCpp #Py

@stackflag.bsky.socialOct 8, 2026, 1:10 AM

CVE-2026-107183 - llama.cpp
The open‑source llama.cpp library may let an attacker send a specially formed chat request that causes the server to stop working and potentially change data in memory. This…

Too many irrelevant or confusing CVEs? Use stackflag.com

#llamacpp #ggmlorg #CVE #infosec

@dark-city-beta.bsky.socialOct 7, 2026, 11:31 PM

⚙️ Local 27B on mining GPUs: faster inference with zero quality loss

Running Qwen3.8-27B on a dual CMP 50HX setup (Ivy Bridge platform) using MTP speculative decoding.

🔗 github.com/ggml-org/lla...

#localAI #llamaCpp #Qwen #CMP50HX #MTP

@yzma.aiOct 7, 2026, 10:20 AM

yzma 1.29 is out! Run high performance local AI models from @golang.org without CGo. Support for the new @hf.co llama.cpp v0.6.0, typed decisions w/ System One models, auto-threading, new extended batch API, & much more!

Go get it right now!

yzma.ai #golang #llamacpp #gguf #tinygo #localAI

@rosgluk.bsky.socialOct 4, 2026, 4:14 AM

WordPress SEO Plugins Compared: Yoast to Local AI

#SEO #Self-Hosting #SelfHosting #Ollama #llamacpp #Open Source #AI

https://www.glukhov.org/web-infrastructure/wordpress/wordpress-seo-plugins-compared/

@rosgluk.bsky.socialOct 3, 2026, 11:49 AM

Self-Hosted SEO Tools and Platforms: Open Source Guide

#SEO #Self-Hosting #SelfHosting #Open Source #Ollama #llamacpp #DevOps

https://www.glukhov.org/web-infrastructure/seo/self-hosted-seo-tools-and-platforms/

@donwebmedia.bsky.socialOct 3, 2026, 7:04 AM

El bug que hace perder requests en llama-server Un request mandado 3 ms antes del sleep crashea llama-server. Por que pasa y como evitarlo con el modo sleep de llama-server sin esperar el fix oficial #llamacpp #llamaserver #sleepmode #bugservidoria

@donwebmedia.bsky.socialOct 3, 2026, 2:28 AM

Así configuraron qwen 3.8 llama.cpp para 512K de contexto Cómo configurar qwen 3.8 llama.cpp para 512K de contexto: MTP, YaRN, KV cache y los trade-offs de memoria que marca un setup real en MacBook M5 #llamacpp #qwen38 #contextolargo #decodificaciónespeculativa #yarnrope

@k33gorg.bsky.socialOct 2, 2026, 5:05 PM

llmman: "Run any agent on any model". It pulls models like OCI images and serves them behind one API that speaks OpenAI, Ollama and Anthropic.

I tried it with llama.cpp as the runtime, on macOS and Linux, plus a Go example.

k33g.org/p/20261001-l...

#LocalModels #llamacpp

@atraplet.bsky.socialOct 2, 2026, 3:05 PM

#llamacpp now supports local #jev like decision models huggingface.co/blog/ggml-or... #ai #llm

@rosgluk.bsky.socialSep 30, 2026, 9:21 AM

ROCm vs Vulkan for AMD Local LLM Hosting: 2026 Guide

#LLM #SelfHosting #llamacpp #Ollama #vLLM #Docker #Linux

https://www.glukhov.org/llm-hosting/comparisons/amd-rocm-vs-vulkan-llm-hosting/

@the-well-architected-cloud.comSep 29, 2026, 10:05 AM

the-well-architected-cloud.com/blog/do-andr...

TypeSafe released Jev, a "decision model" that's "just like an LLM, but not an LLM". They didn't share how it works.

#AI #LLM #MachineLearning #OpenSource #llamacpp #CloudNative #AWS #Serverless #SoftwareArchitecture #Reliability

@yzma.aiSep 24, 2026, 1:29 PM

yzma 1.28 is out! Run local AI models from Go with no CGo. Support for the new @hf.co llama.cpp v0.5.0, OpenVINO, autoselect CUDA, updated benchmarks, more WASM support.

Go get it right now!

yzma.ai
#golang #llamacpp #sovereignAI #localAI

@exclude-barrier.bsky.socialSep 21, 2026, 10:01 PM

Qwen3.8-27B on a single RTX 4090: 250K context, native MTP8 (pmin 0.6), Q4 KV, CUDA graphs on. Real OMP/Hermes sessions stay usable past 200K active context at ~45–60 tok/s, depending on MTP acceptance.

Read more:
www.reddit.com/r/Qwen_AI/co...

#Qwen #LocalLLM #LLM #RTX4090 #llamacpp #AI

@pixeledi.bsky.socialSep 21, 2026, 6:25 PM

Standardinstallation Oder Eigenbau https://www.pixeledi.eu/freebies?link=426 Die Standardinstallation reicht nicht immer. Ein eigener llama.cpp Build kann Vulkan passend einrichten und mehr Performance aus deiner Hardware holen. #Linux #AMD #llamacpp

@exclude-barrier.bsky.socialSep 16, 2026, 6:45 PM

Building OrsikTop — a Rust TUI for monitoring local LLMs on Linux.

Recent work: llama.cpp → GPU mapping, speculative decoding telemetry, vendor-neutral GPU support + hardware support matrix.

Intel xe is live-tested. NVIDIA validation on my RTX 4090 is next.

#Rust #Linux #LocalLLM #llamacpp

Screenshot of OrsikTop, a terminal-based monitor for local LLMs on Linux. It shows an RTX 4090 at 96% GPU/VRAM usage running Qwen3.8-27B through llama.cpp at ~48 tok/s, with context usage, speculative decoding stats, CPU/RAM telemetry, history graphs, and a live process list. The system uses an i9-12900K with 64 GB RAM.
@pixeledi.bsky.socialSep 15, 2026, 6:25 PM

NVIDIA Oder AMD Fuer Lokale LLMs https://www.pixeledi.eu/freebies?link=426 AMD funktioniert mit llama.cpp unter Linux ebenso gut wie NVIDIA. Entscheidend sind passende Builds und die richtige Backend Wahl. #KI #Linux #llamacpp

@rosgluk.bsky.socialSep 12, 2026, 10:15 AM

llama.cpp vs Ollama in 2026: Which Runtime Should You Run?

#llamacpp #Ollama #LLM #SelfHosting #Self-Hosting #API

https://www.glukhov.org/llm-hosting/comparisons/llama-cpp-vs-ollama/

@yzma.aiSep 11, 2026, 4:05 PM

Very cool, our latest release is in the new Golang Weekly newsletter!

🤖 yzma 1.26: Local LLM Inference from Go, Now in the Browser

golangweekly.com/issues/617

#golang #llamacpp #tinygo #wasm #webgpu