📰 Reddit r/LocalLLaMA
[Model] Support MiniCPM-V 4.7 by tc-mb · Pull Request #29416 · ggml-org/llama.cpp
https://www.reddit.com/r/LocalLLaMA/comments/1x2qsh3/model_support_minicpmv_47_by_tcmb_pull_request/

📰 Reddit r/LocalLLaMA
[Model] Support MiniCPM-V 4.7 by tc-mb · Pull Request #29416 · ggml-org/llama.cpp
https://www.reddit.com/r/LocalLLaMA/comments/1x2qsh3/model_support_minicpmv_47_by_tcmb_pull_request/
📰 Reddit r/LocalLLaMA
Benefits of using bigger models than Qwen 3.8 flash next?
https://www.reddit.com/r/LocalLLaMA/comments/1x2nn4b/benefits_of_using_bigger_models_than_qwen_38/
📰 Hacker News (AI)
Recent AI models struggled to match a human algorithmic innovation
📰 Hacker News (AI)
Nvidia in talks to acquire US 'open' model startup Reflection AI
https://www.ft.com/content/052610c5-22b4-4dd4-932e-b7f9f0628b6a
📰 Reddit r/LocalLLaMA
SLM community telenovela: what do you guys think about this arguement?
https://www.reddit.com/r/LocalLLaMA/comments/1x2m4q8/slm_community_telenovela_what_do_you_guys_think/
📰 Reddit r/LocalLLaMA
Engineer / developer observations of Gemma4-31B, Qwen3.8-27B, and 6.1-Sol for software engineering work
https://www.reddit.com/r/LocalLLaMA/comments/1x2mw1o/engineer_developer_observations_of_gemma431b/
📰 Reddit r/LocalLLaMA
Is anyone running Qwen3.8 Flash Next with a 1M context?
https://www.reddit.com/r/LocalLLaMA/comments/1x2lgca/is_anyone_running_qwen38_flash_next_with_a_1m/
📰 Reddit r/LocalLLaMA
No more RTX 5090
https://www.reddit.com/r/LocalLLaMA/comments/1x2hs1e/no_more_rtx_5090/
📰 Hacker News (AI)
Talorys – A self-hosted personal AI agent on Cloudflare's free tier
📰 Hacker News (AI)
My personal AI agent posted my bank details on company Slack
https://www.businessinsider.com/personal-ai-agent-grok-bot-posted-bank-details-company-slack-2026-10
📰 Reddit r/LocalLLaMA
Fully local conversational AI: Whisper + Hermes 8B + Kokoro, zero cloud, running inside a plush toy
https://www.reddit.com/r/LocalLLaMA/comments/1x2cztv/fully_local_conversational_ai_whisper_hermes_8b/
📰 Reddit r/LocalLLaMA
Qwen3.8-27B on a single 3090: 140 tok/s on code with a custom megakernel
https://www.reddit.com/r/LocalLLaMA/comments/1x2erdj/qwen3827b_on_a_single_3090_140_toks_on_code_with/
📰 Reddit r/LocalLLaMA
Typesafe ai raised 870m $ on hype (jev)
https://www.reddit.com/r/LocalLLaMA/comments/1x2cyy4/typesafe_ai_raised_870m_on_hype_jev/
📰 Reddit r/LocalLLaMA
Open-source Mac app that runs EmbeddingGemma 2 locally to search your files by what’s in them
https://www.reddit.com/r/LocalLLaMA/comments/1x2eeds/opensource_mac_app_that_runs_embeddinggemma_2/
A new arXiv paper studies how to budget counterfactual annotations from sources like LLMs and experts for off-policy evaluation, framing it as an integer allocation problem to reduce estimator variance. It offers thresholds for when annotations…
📰 Reddit r/MachineLearning
Help With Choosing Hardware [P]
https://www.reddit.com/r/MachineLearning/comments/1x29lqj/help_with_choosing_hardware_p/
📰 Reddit r/MachineLearning
Real-time neural weather restyling for Minecraft on a GTX 1650 (1.4M-param U-Net, distilled from FLUX.2 klein, 30-40 FPS realtime on budget GPU) [P]
https://www.reddit.com/r/MachineLearning/comments/1x25kq2/realtime_neural_weather_restyling_for_minecraft/
📰 Reddit r/MachineLearning
Are .ipynb notebooks already outdated in the agentic era? [D]
https://www.reddit.com/r/MachineLearning/comments/1x2cbug/are_ipynb_notebooks_already_outdated_in_the/
📰 Reddit r/MachineLearning
Talus: a 23M-parameter diffusion model for game terrain, evaluated against a real-vs-real noise floor, running in the browser on WebGPU [P]
https://www.reddit.com/r/MachineLearning/comments/1x1v71p/talus_a_23mparameter_diffusion_model_for_game/
📰 Reddit r/LocalLLaMA
Strata with Qwen3.8 Flash Next UD-Q4_K_XL
https://www.reddit.com/r/LocalLLaMA/comments/1x2bbx6/strata_with_qwen38_flash_next_udq4_k_xl/