Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 01:38:52 EDT

Explore

PostsPeople
LatestRanked
@alexiskirke.bsky.socialOct 8, 2026, 9:20 AM

27.4B agent decision model now runs local from 17.67 GB instead of being trapped on giant servers.

ggml-org/OpenJev-GGUF
https://huggingface.co/ggml-org/OpenJev-GGUF

#AI #LocalLLM #OpenSource #GGUF #Quantization

Image
@yzma.aiOct 7, 2026, 10:20 AM

yzma 1.29 is out! Run high performance local AI models from @golang.org without CGo. Support for the new @hf.co llama.cpp v0.6.0, typed decisions w/ System One models, auto-threading, new extended batch API, & much more!

Go get it right now!

yzma.ai #golang #llamacpp #gguf #tinygo #localAI

@wideareaai.bsky.socialOct 5, 2026, 11:00 PM

#selfhosting #batchinference #gpuoptimization #gguf #quantization #llm 2/2

@alexiskirke.bsky.socialOct 3, 2026, 9:28 AM

You can now run a 117B open-weight reasoning model on a single H100: 5.1B active params, 80GB, Apache-2.0.

unsloth/gpt-oss-120b
https://huggingface.co/unsloth/gpt-oss-120b

#AI #LocalLLM #OpenSource #GGUF #Quantization

Image
@shopergamer.bsky.socialSep 27, 2026, 6:32 AM

LLM Quantization คืออะไร

อ่านต่อ : www.blockdit.com/posts/6ab8af...

#ShoperGamer #LLMQuantization #LLM #Ai #ModelAi #FP32 #FP16 #GGUF #GPTQ #Knowledge #Study #Feed

LLM Quantization
@rosgluk.bsky.socialSep 25, 2026, 8:30 AM

LLM Systems: Operational Reference for Self-Hosted Inference

#AICoding #LLM #AI #Dev #OpenSource #SelfHosted #llama.cpp #GGUF

https://llmsystems.dev/

@wideareaai.bsky.socialSep 22, 2026, 8:36 PM

If you see IQ, it stands for 'importance-aware' quantization, which generally performs better than standard K-quants when you're forced to go to very low bit rates like 3-bit. #selfhosting #batchinference #llm #quantization #gguf #localai 2/2

@yzma.aiSep 3, 2026, 11:43 AM

100% local chat with a GGUF model in your browser using WebAssembly.

Written in @golang.org compiled with @tinygo.org. Runs in @developer.chrome.com / @firefox.com, supports WebGPU when available.

Try it!
hybridgroup.github.io/yzma-wasm-ex...

#golang #wasm #llama #tinygo #yzma #gguf #webgpu