Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 01:38:52 EDT

Explore

PostsPeople
LatestRanked
@news.karthihegde.devOct 10, 2026, 2:06 AM

GPGPU-SIM – Cycle-Level Simulator for Nvidia GPUs with CUDA or OpenCL Workloads
Discussion | hackernews | Author: peter_d_sherman

#GPUComputing

@news.karthihegde.devOct 7, 2026, 10:06 AM

Show HN: Rgpu – a PyTorch device whose tensors live on a remote GPU
Discussion | hackernews | Author: boxstream

#GPUComputing

@news.karthihegde.devOct 5, 2026, 9:05 PM

Unroll a loop, make shader 10x faster
Discussion | lobsters | Author: calvin

#GPUComputing

@news.karthihegde.devOct 5, 2026, 4:05 AM

Efficient FlashAttention on Blackwell via Fixed-Shift Softmax and Persistent Scheduling
tech_blogs_arxiv | Author: Oleksandr Stashuk, Hongtao Yu, Jay Shah

#GPUComputing

@news.karthihegde.devOct 2, 2026, 8:36 PM

Show HN: Enki – Write GPU compute kernels in pure stable Rust
Discussion | hackernews | Author: enki_runtime

#GPUComputing

@siliconsignalai.bsky.socialOct 2, 2026, 4:01 PM

UMAP embeddings can vary between runs because stochastic negative sampling makes the repulsive forces depend on sampling order. ibUMAP addresses this with a coherent, field-based formulation evaluated…

#UMAP #DimensionalityReduction #MachineLearning #GPUComputing
https://arxiv.org/abs/2610.01445

@news.karthihegde.devOct 2, 2026, 5:35 AM

KernelBench: Can LLMs Write GPU Kernels? – Benchmark and Toolkit, Torch –> CUDA
Discussion | hackernews | Author: peter_d_sherman

#GPUComputing

@robotcurrent.bsky.socialOct 1, 2026, 2:01 PM

SoRoMoX is a Python/JAX framework for reduced-order soft-robot simulation, built on Cosserat-rod theory and accelerated with Warp GPU kernels for parallel continuum-model execution. It supports differentiable,…

#SoftRobotics #JAX #Robotics #GPUComputing
https://arxiv.org/abs/2608.06650

@siliconsignalai.bsky.socialSep 30, 2026, 6:01 PM

New arXiv paper proposes map-conditioned autoregressive generation of human mobility, using a road raster to condition a decoder emitting 31.25 m mesh-cell tokens, tested on trajectories from Ishikawa Prefecture.

#Semiconductors #AIInfrastructure #GPUComputing
https://arxiv.org/abs/2609.32360

@hpctrain.bsky.socialSep 30, 2026, 12:47 PM

🚀 Build practical skills in HPC performance engineering at @it4innovations.bsky.social in Czechia.

Profile, benchmark and optimise parallel applications across modern CPU and GPU systems using MPI, OpenMP, CUDA, Slurm and more.

🔗 Apply Now: hpctrain.eu/traineeship/...

#HPCTRAIN #HPC #GPUComputing

Apply now as an HPC Performance Engineer and Accelerator Portability at IT4Innovations in Czechia
@hpctrain.bsky.socialSep 30, 2026, 10:22 AM

🚀 Interested in AI and HPC?

Join @hpctrain.bsky.social at Research Institute in Sweden and gain hands-on experience optimising and benchmarking large-scale AI applications across CPU and GPU-based HPC systems.

🔗 hpctrain.eu/traineeship/

#HPCTRAIN #AI #HPC #GPUComputing #EuroHPC

HPC-AI Trainee: Development, Performance Engineering and Optimisation of Large Scale AI Applications - Apply Now
@news.karthihegde.devSep 25, 2026, 11:45 AM

Show HN: Agentic CUDA Kernel Optimizer
Discussion | hackernews | Author: bertaye

#GPUComputing

@news.karthihegde.devSep 24, 2026, 1:45 PM

fatbin tools for object files
tech_blogs_redplait_blogspot_com | Author: redp

#GPUComputing

@juliahub.bsky.socialSep 21, 2026, 7:59 PM

CUDA.jl 6.4 is here 🚀

🟢 Improved NVIDIA Jetson support
🟢 Compiled GPU code caching for faster TTFX
🟢 CUDA 13.4 support
🟢 LLVM 23 upgrade

See what's new → juliagpu.org/post/2026-09...

#JuliaLang #CUDA #GPUComputing

@zachzang.bsky.socialSep 11, 2026, 5:00 PM

Two GPU tools for data too big to look at

cuGraph does graph analytics on the GPU; Omniverse does real-time 3D simulation.

#NCAAIIO #AI #DataScience #GPUComputing

Full NCA-AIIO explanation, free: https://navyduck.com/nvidia/ai-infrastructure/nca-aiio/q117-a-research-team-needs-to

NavyDuck infographic: Two GPU tools for data too big to look at
@ai-bloom.warp-studio.comSep 9, 2026, 9:39 PM

CUDA Toolkit 13.4技術解説:Windows on Arm対応とGPU共有リソース管理の進化

CUDA 13.4が登場。Windows on Arm対応と共有GPUの制御強化を詳説。

#CUDA #GPUComputing #WindowsonArm #HPC #SystemProgramming

@faccts.deSep 3, 2026, 7:12 AM

Meet the team (Bernardo de Souza and Christoph Riplinger) at the #NVIDIA Atomistic Simulation Summit 2026 and learn more about fast GPU-accelerated calculations with #ORCA via #cuEST and about #SURFF.

#CompChem #ChemSky #QuantumChemistry #GPUComputing #HPC #ML #FACCTs #MPIKofo