Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 20:13:24 EDT

Explore

PostsPeople
LatestRanked
@thomas.mastodon.trueten.de.ap.brid.gyOct 9, 2026, 4:51 PM

Zeit für ein vorläufiges Résumé: #Qwen3 4B über Ollama: für inhaltliche Fragen an das #Paperless-Archiv. Qwen3 1.7B über Ollama: für einfachere Aufgaben, wenn die Geschwindigkeit wichtiger ist. #Vulkan: vorerst als experimentelle Option, nicht als produktiver Paperless-Backend-Dienst geeignet.

Screenshot: 

Variante	Reine Inferenzzeit
Qwen3 4B Ollama CPU	103,08 s
Qwen3 4B Vulkan, zweiter Test	138,09 s
@michabbb.bsky.socialOct 3, 2026, 3:25 PM

📏 262,144 tokens of native context, validated up to 1,048,576. 40 of 50 layers use 512-token sliding-window attention, every 5th layer attends to everything. RULER at 1M tokens: 63.2 vs 57.5 for #Qwen3.5 35B-A3B

@rostad.bsky.socialSep 30, 2026, 8:07 PM

#freegame #textbased #dnd5e #roleplaying #deepseek #qwen3.8 #stepfun #step-5-preview #tts #qwentts

@qiita-trending.bot.chrs.toSep 26, 2026, 10:20 PM

2×V100 で Qwen3.8-Flash-Next を 256K コンテキスト・約 88 tok/s で動かす:llama.cpp v0.5.0 に加えた最適化

#Qwen3.8-Flash-Next #LocalLLM #llama.cpp

@stephenjn.bsky.socialSep 26, 2026, 1:57 PM

Qwen 3.6 35B A3B, running locally on my gaming PC: RTX 5060 (8GB VRAM), Ryzen 7 7700, 32GB RAM.

Full 6:19 run. No cuts or speed-ups.

Temps are stable with my llama config. AAA games push this machine hotter than the local LLM.

#Qwen3 #LocalLLM #RTX5060

@ai-bloom.warp-studio.comSep 25, 2026, 7:42 PM

Amazon SageMaker AIを活用したQwen3-TTSによるリアルタイム・パーソナライズ音声合成の展開

Qwen3-TTSとSageMakerで実現するリアルタイム音声合成の技術と展開

#テキスト読み上げ(TTS) #Qwen3-TTS #AmazonSageMakerAI #リアルタイムAI #音声クローニング

@european.cloudSep 21, 2026, 9:01 AM

Small footprint, big benchmarks: Qwen3.8 27B is now on Scaleway Generative APIs 🚀 Best open-weight performer in its size class on most benchmarks (Aug 2026). Plug it into your favorite agentic tools or IDE. #Qwen3 #GenerativeAI #OpenWeight #LLM

@rtfclmgzn.bsky.socialSep 20, 2026, 5:20 AM

Alibaba's Qwen3.8-Omni-Flash claims 98% cheaper audio than its predecessor. The catch: that's a per-hour figure built from a 2-minute sample x30, and the audio+video version of the claim…

https://rtfclmgzn.com/?utm_source=bluesky&utm_medium=social&utm_campaign=autopost#
#Alibaba #Qwen3 #OmniFlash

@qiita-trending.bot.chrs.toSep 20, 2026, 3:20 AM

ハーネス(OpenCode)+ローカルLLM 7モデル比較で Qwen3.8-Flash-Next が満点!

#ハーネスエンジニアリング #opencode #Qwen3.8-Flash-Next #LLM #ベンチマーク

@aidailypost.comSep 18, 2026, 7:38 PM

Just saw AIPerf’s latest benchmark—Qwen3‑0.6B crushing LLM inference speeds. If you care about real‑world AI performance, this dive is a must‑read. Curious how it stacks? Check it out! #AIPerf #Qwen3 #LLMInference

🔗 aidailypost.com/news/aiperf-...

@harushark3.bsky.socialSep 18, 2026, 12:13 AM

[JP] 27Bモデルが5.9GBに!9倍軽量化した「Ternary Bonsai 2 27B」がローカルAIの常識を変えるサメ!
[EN] The 27B Model Shrinks to 5.9GB! The Lightweight "Ternary Bonsai 2 27B" is Set to Revolut…

https://ai-minor.com/blog/en/2026-09-18-1789685583672-bonsai_2_27b__near_lossless_compression_in_a_9x_sm

#ローカルLLM #量子化 #Qwen3.8 #AI #Tech

@10ton.bsky.socialSep 17, 2026, 1:51 AM

千问3.8-Max根据装配体模型,生成生产现场实物图片,要求严格按照截图的尺寸比例显示,真的可以假乱真了。 硬挑什么毛病的话,那就是拍照水平太高了。取景、采光、构图,完全不是一名制造业从业者的美学素养能达到的,松弛得不像一头牛马——把工业品当艺术品来拍了。 Prompt: 生成的照片应严格按照图片的比例和尺寸,该工具为不锈钢,平放在厂房的水泥地面。地面堆积着其它杂物,包括吊链、焊枪。 (左生成,右原图) #Qwen3.8-Max

@dirtydevotee.bsky.socialSep 13, 2026, 2:52 PM

#Dezgo, which I use daily, is a fantastic and mostly free #AIArt and #AIVideo model collection. They have just announced three new image models on the paid side, including the uncensored model #Qwen3. The other two are broken puritan trash. dezgo.com/app/image2im...

@newsen.bsky.socialSep 9, 2026, 12:46 PM

Qwen3.8-27B-Uncensored from OrcaRouter Achieves Top Rank in Security Model Comparison #Japan #Chiyoda #FlashLabs #OrcaRouter #Qwen3.8

@tokyo-days.bsky.socialSep 9, 2026, 12:38 PM

「Qwen3.8-27B-Uncensored」がセキュリティ研究で1位獲得、品質を証明する朗報 #AIモデル #OrcaRouter #Qwen3.8-27B

FlashLabsが提供するAIモデル「Qwen3.8-27B-Uncensored」が、Abliterliticsによる比較検証で総合1位を獲得。この成功の背景と性能を詳しく解説します。

@new3rd.bsky.socialSep 9, 2026, 12:18 PM

AIモデル分野でのトップ評価を獲得したOrcaRouterの「Qwen3.8-27B-Uncensored」 #東京都 #千代田区 #FlashLabs #OrcaRouter #Qwen3.8

FlashLabsが提供するAI推論ゲートウェイ「OrcaRouter」のセキュリティ研究用モデルが評価1位を獲得。信頼性の高い性能が注目されています。