Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-11 02:37:10 EDT

Explore

PostsPeople
LatestRanked
@hgpu.bsky.socialSep 27, 2026, 10:35 PM

Microarchitectural Memory Bandwidth Saturation, KV-Cache Paging Dynamics, and Time-to-First-Token Latency: A Comparative Benchmark of vLLM, TensorRT-LLM, and FlashAttention-3 on NVIDIA Hopper H100 versus AMD Instinct MI300X

#CUDA #ROCm #AMD #vLLM #Benchmarking #Performance

hgpu.org?p=31273

@daaronr.bsky.socialSep 26, 2026, 11:50 PM

But can AI jazz?

A quick experiment: ai-jazz-tune-trial.netlify.app

Have a listen and vote: Which model's got more swing? #jazz #ai #benchmarking #music

@tmlr-pub.bsky.socialSep 26, 2026, 4:20 AM

Benchmarking Transfer Learning: From Simple Baselines to Combined Scorers for Transferability Est...

Levy Chaves, Claudio Mayrink Verdun, Eduardo Valle, Sandra Avila

Action editor: Dmitry Kangin

https://openreview.net/forum?id=3i2ZRk8GDN

#imagenet #transferability #benchmarking

@gbti.bsky.socialSep 25, 2026, 5:05 PM

Shared on the GBTI Network: "AI plays Age of Empires II"

#gaming #ai #ageofempires #grok #gemini #openai #anthropic #benchmarking

www.youtube.com/watch?v=ZBdA...

@tmlr-pub.bsky.socialSep 23, 2026, 8:20 PM

RAWDet-7: A Multi-Scenario Benchmark for Object Detection and Description on Quantized RAW Images

Mishal Fatima, Shashank Agnihotri, Kanchana Vaishnavi Gandikota et al.

Action editor: Tatsuya Harada

https://openreview.net/forum?id=UHTJrsYieo

#benchmarking #benchmark #srgb

@gbti.bsky.socialSep 22, 2026, 11:58 PM

In this report, frontier models are benchmarked on their ability to play StarCraft: Brood War:

gbti.network/shares/gbtil... #gaming #BroodWar #codex #claudes #gbti #starcraft #ai #benchmarking #aigaming #claude #fable #grok #gpt #codex

@adaptiveus.bsky.socialSep 22, 2026, 12:18 PM

#TechniquesforTuesday
Benchmarking helps Business Analysts answer that question by comparing processes, performance, and practices against competitors, industry peers, or best-in-class organizations.
Read more...
www.adaptiveus.com/blog/busines...

#benchmarking #adaptiveus #batechniques

@yzhums.bsky.socialSep 21, 2026, 11:57 PM

Improving email security outcomes with real-world Microsoft Defender insights
www.microsoft.com/en-us/securi...

#email
#security
#Microsoft
#Benchmarking

@t8ngy.bsky.socialSep 19, 2026, 2:27 PM

#تكنولوجيا #AI #Benchmarking #TechNews

@thedailytechfeed.comSep 19, 2026, 1:25 PM

Vals wants you to trust AI benchmarks again: unseen tests, ethics, sector-specific rigor. #AI #Benchmarking #Vals #A16z #ModelEvaluation #AITrust https://thedailytechfeed.com/vals-aims-to-set-new-benchmark-standards-for-ai-validation/

@deverauxdev.bsky.socialSep 19, 2026, 1:35 AM

My previous multi-core estimates were inflated by a compiler artifact, so I built a hostile, self-policing benchmark to find the true physical floor of my code.
#RustLang #SystemsProgramming #Benchmarking #BuildinPublic

@stridingtech.bsky.socialSep 18, 2026, 6:02 AM

Traditional benchmarks falter against heterogeneous silicon. Our investigation details how advanced telemetry is essential to accurately decode chip performance complexities. #Semiconductors #Benchmarking https://stridingtech.com/archives/6808

@tmlr-pub.bsky.socialSep 17, 2026, 4:25 PM

New #Reproducibility Certification:

Benchmarking Tabular Foundation Models for Conditional Density Estimation in Regression

Rafael Izbicki, Pedro L. C. Rodrigues

https://openreview.net/forum?id=KWsWHpp5Do

#benchmark #benchmarking #prediction

@go-euc.bsky.socialSep 17, 2026, 10:01 AM

At GO-EUC, there’s no one-size-fits-all benchmark.

Depending on what we’re testing, we use solutions like #LoadGen, #LoginEnterprise, and most importantly, #OBUX to capture the data that actually matters.

www.go-euc.com/insight-in-t...

#EUC #Benchmarking #GOEUC

@mlcommons.orgSep 16, 2026, 2:53 PM

MLPerf Inference v6.1 results are live.

Record 30 organizations, two new benchmarks (End-to-End RAG + Edge Agentic Inference), and performance gains accelerating — Deepseek R1 up 5.7X in a year, VLM up 2.99X in six months.

mlcommons.org/2026/09/mlpe...

#AI #Benchmarking #AgenticAI #RAG #EdgeAI

@werewolfpack.bsky.socialSep 15, 2026, 7:30 PM

Running Super Pi 32M on a 2004 Toshiba Satellite L10 (Celeron M 370)
youtube.com/shorts/E04Sl...

#technews #news #pc #computer #vintagehardware #puperpi #2004laptop
#toshibasatellitel10-108 #YouTube #retrohardware #benchmarking #benchmark

@informaq.bsky.socialSep 15, 2026, 9:09 AM

New benchmark study reveals practical method for comparing quantum computers by measuring computational success rates and execution speed, indicating ~100,000-fold performance gains needed before scientific viability.

#QuantumComputing #Benchmarking #News

@axompi.bsky.socialSep 9, 2026, 1:29 PM

* Do you have any publication on robotic grasping and manipulation, benchmarking, reproducibility, or AI for robotics?

* Is the paper accepted at #IROS2026?

* Are you only attending IROS but have other relevant works?

#Robotics #Grasping #Manipulation
#Benchmarking #Reproducibility #AI #Humanoid

@tmlr-pub.bsky.socialSep 5, 2026, 8:20 AM

GEO-Bench-2: From Performance to Capability, Rethinking Evaluation in Geospatial AI

Naomi Simumba, Nils Lehmann, Paolo Fraccaro et al.

Action editor: Frederic Sala

https://openreview.net/forum?id=NPf175jnP1

#benchmarking #terramind #imagenet