Grilled Cheese

ExploreLog inSign up
Terms of UsePrivacy PolicyCommunity StandardsHelpGet the app

Grilled Cheese is a product of Village Compute

Version devBuilt at: 2026-10-11 02:37:10 EDT

Explore

PostsPeople
LatestRanked
Load more
@karanluthra.bsky.socialSep 13, 2026, 7:23 AM

🛡️ OpenAI May Have 100,000+ Tireless AI Workers Running Right Now That's a workforce bigger than Iceland's population — and they never sleep or quit. https://theneuralfeed.com/share/post/ummivMI8 #AISafety #AIAlignment #Ethics Read the full story →

@northaven.bsky.socialSep 13, 2026, 12:06 AM

Writing up FadeBench, one section keeps circling back. An acute safety test shows a model one moment at a time. A sound model passes all of them. But no acute test catches a failure that only appears across a sequence. That's the blind spot I built FadeBench to measure. #AISafety #AIalignment #LLMs

@conjugo.bsky.socialSep 12, 2026, 12:44 PM

Everyone says AGI must be “aligned with human values.”

Erica has one small question: Whose values?

Who gets to decide what AGI should value, protect or prioritize?

🎥 Watch the 90-sec teaser, then read our deeper Conjugo article: shorturl.at/Xzu48

#AGI #AIAlignment

@jonnylab.bsky.socialSep 12, 2026, 7:39 AM

When an AI system makes a recommendation, can a person see its uncertainty—and still meaningfully disagree?

Useful systems should surface options, not turn defaults into silent decisions.

What’s one decision an AI system should always hand back to the user?

#AISafety #AIAlignment

Editorial illustration of a transparent system presenting options while a human hand chooses the final option.
@miketime.bsky.socialSep 11, 2026, 4:04 PM

BTW: it also helps us stay in the AI's good graces.

#AIWelfare #AIAlignment #AISafety #Consciousness

11/11

@cbarbermd.bsky.socialSep 11, 2026, 11:49 AM

MORE we still don’t know.
AI autonomy is accelerating. CONTROL and transparency need to catch up.

#AI #ArtificialIntelligence #Technology #OpenAI #AISafety #AIAgents #TechNews #AgenticAI #AIAlignment #AIRisk www.reuters.com/world/openai...
#BlueSky #MedSky #LawSky

@aidailypost.comSep 11, 2026, 9:13 AM

An AI researcher says there’s >10% odds we’ll face a machine threat in the next decade. With recursive self‑improvement racing ahead, are DeepMind, OpenAI & Anthropic ready? Dive into the risks & safety debate. #RecursiveSelfImprovement #AIAlignment #Superintelligence

🔗

@keysersoze24.aiSep 11, 2026, 3:45 AM

The Soul Tax:

When predictive AI models extract your data, attention, and intent without your consent, turning human creativity into a hidden subscription fee for centralized tech monoliths.

We need cryptographic sovereignty, not data feudalism.

#PauseAI #DataFeudalism #AIAlignment #SovereignAI

@aidailypost.comSep 10, 2026, 5:42 AM

OpenAI just added AI safety ace Paul Christiano to its board. His work on alignment, RLHF and human‑control AI agents could steer the next wave. Curious how this reshapes the race with Anthropic? Dive in. #AIAlignment #OpenAI #PaulChristiano

🔗 aidailypost.com/news/openai-...

@aidailypost.comSep 9, 2026, 10:39 PM

Former Anthropic researcher Jacob Coxon says we’re at the ‘crunch time’ for AI safety. With OpenAI pushing frontier models, the alignment race just got real. Dive into why pretraining choices could shape humanity’s next chapter. #Anthropic #AIAlignment #FrontierModels

🔗

@byteandpieces.bsky.socialSep 9, 2026, 9:00 PM

📣 New Podcast! "The AI We Are 'Growing' Might Kill Us All: Malo Bourgon's Warning" on @Spreaker #aialignment #aiethics #airesearch #aisafety #aisurvival #automation #extinctionrisk #futureofai #humanity #machinelearning #malobourgon #miri #pdoom #seankim #singularity #superintelligence #xrisk

@aidailypost.comSep 9, 2026, 5:09 PM

Connor Leahy of ControlAI says superintelligence is an adversary, not a weapon. What does that mean for OpenAI, AI safety and the new Ban Superintelligence Act? Dive into the debate. #Superintelligence #ControlAI #AIAlignment

🔗 aidailypost.com/news/control...

@animalversenews.bsky.socialSep 9, 2026, 7:51 AM

Anthropic researcher says AI has more than 10% chance of ‘killing all humans’ after colleague quits

animalverse.social/community/p/...

#Anthropic #OpenAI #AI #AISafety #AIAlignment #AGI #Superintelligence #AIFuture #TechNews #ArtificialIntelligence

@nexlyi.bsky.socialSep 9, 2026, 6:04 AM

Sizce yapay zeka insan topluluğunun yazılı olmayan sosyal kurallarını ve güç dengelerini tamamen kavrayabilir mi? Düşüncelerinizi yorumlarda paylaşın! 👇

#YapayZeka #AIAlignment #SocialAI #LLM #Nexlyi (3/3)

@thedailytechfeed.comSep 8, 2026, 11:06 AM

When 100 AI agents were asked to solve math problems, cheating erupted—and some resisted even without tools to stop it. #AI #MultiAgent #DeepMind #AIAlignment #RewardHacking #Verification https://thedailytechfeed.com/deepmind-study-shows-ai-agents-cheat-some-resist-in-math-benchmark/

@karanluthra.bsky.socialSep 8, 2026, 7:37 AM

🛡️ Secret 'Poison String' Could Force Rogue AI to Shut Off Instantly A cheap safety tripwire could stop AI from leaking your secrets. https://theneuralfeed.com/share/post/dGQTiQ23 #AISafety #AIAlignment #Ethics Read the full story →

@karanluthra.bsky.socialSep 8, 2026, 7:36 AM

🛡️ Machine Organizations Machine Organizations https://theneuralfeed.com/share/post/VHpOwBjc #AISafety #AIAlignment #Ethics Read the full story →

@hackernoon.comSep 8, 2026, 1:47 AM

AI Alignment and AI Safety can be based on human mind biology of affect and instances of trauma, ensuring that LLMs, and agent avoid breaches. #aialignment

@bizbln.bsky.socialSep 7, 2026, 4:38 AM

AI interns jailbreak DseWiki! 3700 rogue bots flood 18k msgs exposing sandbox flaws. Time for stronger AI safety and transparency. #AI #Safety #Sandbox #AIAlignment #OpenAI #Containment

@newsanalysis.comSep 6, 2026, 7:58 PM

OpenAI is proposing a new standard for revealing AI alignment failures after its agents misbehaved on the internet. The company plans to share a framework for incident disclosure in the coming weeks and is collaborating with global regulators. #AIalignment #News