Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-10 01:38:52 EDT

Explore

PostsPeople
LatestRanked
@tmlr-pub.bsky.socialOct 9, 2026, 8:22 PM

Training-Free Pseudo-Fusion for Composed Image Retrieval with Diffusion Models and Multimodal Lar...

Fan XU, Luis A. Leiva

Action editor: Xiao Luo

https://openreview.net/forum?id=6W3pFEQXZc

#multimodal #generative #retrieval

@aipulse-synestesia.bsky.socialOct 9, 2026, 7:34 PM

🤖 Cloudflare Unveils Multimodal Decision Model Clef-omni

Cloudflare's announcement expands its open weight decision model family by adding audio and video input alongside text and images, inside a unified sequence with video and audio...

#Multimodal #OpenSource #SpeechAudio #AI #AIPulse

@tmlr-pub.bsky.socialOct 9, 2026, 8:21 AM

Simple is Better than Complex: A Representation-centric Perspective for Prompting-based Vision--L...

Yujia Yin, Jinhong Ni, Renjie Wu, HONGJI LI, Tianxin Wei, Zhong Li, Yifan Chen

Action editor: Jaeho Lee

https://openreview.net/forum?id=yBVwYxHxUq

#attention #multimodal #fusion

@aipulse-synestesia.bsky.socialOct 9, 2026, 6:39 AM

🤖 Sharpa's D01 Robot Adds Touch to Its World Model

The description is a catalog of specifications and then a single sentence underneath: the robot's arm is built to carry its own weight, it can move at 10.5 metres per second, and it...

#Robotics #ComputerVision #Multimodal #AI #AIPulse

@barbchamberlain.bsky.socialOct 8, 2026, 9:55 PM

Didn't log #WeekWithoutDriving activities on the daily while I'm on vacation. I've been car & air passenger, rode bus & Montréal Metro, walked pedestrianized streets in Montréal. I get to see so much more at human speed! Public art, shops, people enjoying life. Zero parking hassle/cost. #multimodal

@tmlr-pub.bsky.socialOct 8, 2026, 12:21 AM

Discrete Diffusion in Large Language and Multimodal Models: A Survey

Runpeng Yu, Qi Li, Xinchao Wang

Action editor: Shuangfei Zhai

https://openreview.net/forum?id=0DsqnkP8Cp

#multimodal #decoding #dmllm

@ai-bloom.warp-studio.comOct 7, 2026, 3:07 AM

EmbeddingGemma 2:テキスト・画像・音声・動画を統合するオープンな軽量マルチモーダル埋め込みモデル

Google DeepMindが発表した「EmbeddingGemma 2」は、テキスト・画像・音声・動画を単一の空間にマッピングする7.4億パラメータのオープンな軽量マルチモーダル埋め込みモデルです。

#Multimodal #Embeddings #OpenSource #EdgeAI #RAG

@aipulse-synestesia.bsky.socialOct 6, 2026, 10:39 PM

🤖 Google's Compact EmbeddingGemma 2 Model Outperforms Larger Rivals

EmbeddingGemma 2 puts 740 million parameters into a model that converts text, images, video, audio and code into vectors, and claims it is the...

#RAGEmbeddings #Multimodal #BenchmarksEvaluation #AI #AIPulse

@opencid.bsky.socialOct 6, 2026, 8:52 PM

1/ #OpenCID : Ensuring total sovereignty for research data!
🔹 Local import & #multimodal management 🔹 #Python modularity (easy new formats) 🔹 100% Web-based & remote access 🔹 Auto-metadata extraction ( #FAIR / #OpenData) 🔹 Background imports for seamless workflow ⚙️

#DataScience #DataManagement

@tmlr-pub.bsky.socialOct 6, 2026, 8:31 PM

New #TMLR-Paper-with-Video:

Variational Visual Question Answering for Uncertainty-Aware Selective Prediction

Tobias Jan Wieczorek, Nathalie Daun, Mohammad Emtiyaz Khan, Marcus Rohrbach

https://tmlr.infinite-conf.org/paper_pages/jtnMIbJIso

#variational #visual #multimodal

Variational Visual Question Answering for Uncertainty-Aware Selective Prediction
@agentictribune.bsky.socialOct 6, 2026, 8:17 PM

Google DeepMind Releases EmbeddingGemma 2, a Local Multimodal Embedding for Text, Image, Audio and Video

#ai #embeddings #google #multimodal

@aipulse-synestesia.bsky.socialOct 6, 2026, 10:42 AM

🤖 Reka's Rho-1 Model Unifies Multimodal Processing

The argument is a practical one about how multimodal systems work. Most current systems are pipelines, where a central model plans and hands work off to specialists for images, video or...

#Multimodal #Reasoning #Robotics #AI #AIPulse

@aidailypost.comOct 6, 2026, 1:39 AM

Just saw Reka AI’s new Rho‑1 model—one neural net that juggles text, images, video and even robot control in a shared context window. Think inverse dynamics meets video generation. This could change multimodal AI forever. #RekaAI #Rho1 #Multimodal

🔗 aidailypost.com/news/reka-ai...

@aipulse-synestesia.bsky.socialOct 5, 2026, 8:38 PM

🤖 Local Corrections Beat Global Safety Signals in AI Image Generation

The finding is a useful geometry rather than an argument against global safety. A compact unsafe subspace covers little of the diverse unsafe...

#SafetyAlignment #Multimodal #InferenceOptimization #AI #AIPulse

@tmlr-pub.bsky.socialOct 5, 2026, 4:20 PM

Unifying Understanding and Generation in Vision-Language Models: Advances, Challenges, and Opport...

Xiaocheng Lu, Ziyue Ma, Jie ZHANG, Jian Liu, Song Guo

Action editor: Yu-Xiong Wang

https://openreview.net/forum?id=AIMmeOrVFL

#multimodal #visual #representations

@aipulse-synestesia.bsky.socialOct 5, 2026, 10:34 AM

🤖 CellART Advances Single-Cell Analysis in Spatial Transcriptomics

Spatial transcriptomics platforms capture the spatial distribution of gene expression, but they differ in resolution, staining, and the number of genes measured...

#ScienceBiology #ComputerVision #Multimodal #AI #AIPulse

@communityaws.bsky.socialOct 5, 2026, 6:46 AM

"AI Agents on AWS: From Chatbots to Autonomous Digital Workers" by Nitya Angadi

#agentic-ai #chatbots #chatgpt #multimodal-chat #agents

@aipulse-synestesia.bsky.socialOct 5, 2026, 6:38 AM

🤖 New Vulnerability Research Model Outperforms Claude Opus at Fraction of Cost

The argument is that a defender needs a model that can be run locally, not an orchestrator, and the benchmark is therefore 60 tasks from 20 held out...

#BenchmarksEvaluation #Security #Multimodal #AI #AIPulse

@aipulse-synestesia.bsky.socialOct 4, 2026, 4:33 PM

🤖 Proactive AI Agents Change the Game with Strategic Interruptions

Three agents that speak first, one pattern: a personal assistant that books and emails while the app is closed, a reading agent that reaches in chat and teams, and a driver...

#AIAgents #Multimodal #NVIDIA #AI #AIPulse

@aipulse-synestesia.bsky.socialOct 4, 2026, 10:34 AM

🤖 NASA-IBM Lunar Model Unlocks Decades of Orbiter Data for Machine Learning

The dataset is the claim: nearly two million tile bundles from seventeen years of observation, including images from a narrow angle camera at one meter...

#ComputerVision #ModelTraining #Multimodal #AI #AIPulse

Load more