π° Reddit r/LocalLLaMA
microsoft doubles down on local ai with nvidia, but with a cost
https://www.reddit.com/r/LocalLLaMA/comments/1x280nz/microsoft_doubles_down_on_local_ai_with_nvidia/

π° Reddit r/LocalLLaMA
microsoft doubles down on local ai with nvidia, but with a cost
https://www.reddit.com/r/LocalLLaMA/comments/1x280nz/microsoft_doubles_down_on_local_ai_with_nvidia/
π° Reddit r/LocalLLaMA
What are people doing with decisions Jev-like models?
https://www.reddit.com/r/LocalLLaMA/comments/1x29q4a/what_are_people_doing_with_decisions_jevlike/
π° Reddit r/LocalLLaMA
want me to get you some?π€£
https://www.reddit.com/r/LocalLLaMA/comments/1x2aqtp/want_me_to_get_you_some/
π° Reddit r/LocalLLaMA
NVIDIA reportedly discontinuing RTX 5090, GB202 GPUs to be reserved for RTX PRO series
https://www.reddit.com/r/LocalLLaMA/comments/1x2bt6r/nvidia_reportedly_discontinuing_rtx_5090_gb202/
π° Reddit r/LocalLLaMA
big or small?
https://www.reddit.com/r/LocalLLaMA/comments/1x29hx8/big_or_small/
π° Hacker News (AI)
LGTM (Looks Good to Me) β Claude Opus 5.5 Music Video
π° Hacker News (AI)
Show HN: Let your AI agents paint big arrows, boxes and text on your screen
π° Hacker News (AI)
Anthropic AI model submits false tip on unsolved Philly murder, police say
https://www.nbcphiladelphia.com/news/local/anthropic-ai-model-submits-false-tip-on-unsolved-philly-murder-police-say/4477051/
New work recasts MIP presolve as autoregressive sequence generation, reframing action order rather than just parameters. ORDO reports zero-shot speedups on unseen domains.
Citrix Patches Critical NetScaler Flaw That Could Enable RCE in SAML Deployments
Citrix has released patches for yet another critical security flaw impacting NetScaler ADC and NetScaler Gateway that could result in remote code execution or denial-of-service (DoS) under certaiβ¦
#hackernews #ml #news
π° Reddit r/LocalLLaMA
What does LocalLLaMA think of REA?
https://www.reddit.com/r/LocalLLaMA/comments/1x212iv/what_does_localllama_think_of_rea/
π° Reddit r/LocalLLaMA
SlopSoup TV. A live, never-ending 24/7 pixel-art TV network inspired by old-school late night Adult Swim. No human makes any of it.
https://www.reddit.com/r/LocalLLaMA/comments/1x20tgs/slopsoup_tv_a_live_neverending_247_pixelart_tv/
π° Reddit r/LocalLLaMA
Qwen3.8-27B UD-IQ4_XS Heretic + MTP on a 16 GB card with 55-68 tok/s (24gb and 12gb versions available too)
https://www.reddit.com/r/LocalLLaMA/comments/1x1zhnx/qwen3827b_udiq4_xs_heretic_mtp_on_a_16_gb_card/
π° Reddit r/LocalLLaMA
Basalt: Flash-Next at 665 tok/s structured, 354 prose on a 5090 + 5060 Ti (2.6x Strata)
https://www.reddit.com/r/LocalLLaMA/comments/1x223ai/basalt_flashnext_at_665_toks_structured_354_prose/
π° Reddit r/LocalLLaMA
Qwen3.8 27b with 200K ctx + MTP on 12GB Ampere Cards
https://www.reddit.com/r/LocalLLaMA/comments/1x1xt1x/qwen38_27b_with_200k_ctx_mtp_on_12gb_ampere_cards/
How identity and permissions become the blast-radius boundary for LLMs scworld.com/perspective/... via SCMagazine
#Identity #Application #security #AI #ML #LLMs
π° Reddit r/MachineLearning
I built ALHR: A tree based sparse attention system that achieves sub-quadratic inference while retaining accuracy. [P]
https://www.reddit.com/r/MachineLearning/comments/1x1lem3/i_built_alhr_a_tree_based_sparse_attention_system/