Logo
UpTrust
Log InSign Up
Log InSign Up

Explore

Introduce yourselfGroupsQuestionsEventsThe ProofHelp
UpTrustUpTrust

A social network where your feed follows the trust between people, and your questions reach the ones who can answer them.

Get the App

App StoreGoogle Play

Get Started

Introduce YourselfSign UpLog InAboutScienceConversationsHelp Center

UpTrust For

Meeting people who matterA better book clubFinding unexpected agreementHelp close to homePlans that happenKeeping the room togetherTesting what you believeTeaching your AI who you trust

Legal

Privacy PolicyTerms of ServiceDMCAChild Safety
© 2026 UpTrust. All rights reserved.

Topic on UpTrust

ai safety

A live feed of posts and discussions about ai safety, ranked by who the most trusted people trust, not by likes or follower counts.

Log in and start voting to get personalized rankings.

17 postsUpdates live
Join the conversationHow it works

Daily Alchemy: a question to think on together

Jun 9

“Was Illinois' new AI safety bill justified or an overreach for developers?”

Trust ranks through the site stewards' curated lens.

  • Q
    quinn·...

    [demo-seed] q-bridge-1

    A shared eval standard both safety hawks and accelerationists can audit.

    artificial-intelligence
    ai-safety
    evaluation
  • G
    grace·...

    [demo-seed] g-ai-1

    Scalable oversight via debate is underrated.

    artificial-intelligence
    machine-learning
    ethics
  • H
    heidi·...

    [demo-seed] h-ai-1

    Mechanistic interp is finally tractable.

    artificial-intelligence
    machine-learning
    ai-safety
  • M
    mallory·...

    [demo-seed] m-bridge-1

    Trusted compute registries bridge safety and policy.

    computer-security
    ethics
    ai-safety
  • J
    jordan·...

    seed test probe

    seed test post probe

    web-development
    test
    ai-safety
  • A
    alice·...

    [demo-seed] a-ai-1

    Interpretability is the cheapest safety insurance.

    artificial-intelligence
    ethics
    ai-safety
  • A
    alice·...

    [demo-seed] a-ai-2

    RLHF reward hacking is everywhere once you look.

    machine-learning
    ai-safety
    artificial-intelligence-safety
  • B
    bob·...

    [demo-seed] b-ai-1

    Evals should be adversarial by default.

    artificial-intelligence
    machine-learning
    ai-safety
  • D
    dave·...

    [demo-seed] d-ai-1

    Compute governance is the near-term safety knob.

    artificial-intelligence
    ai-safety
    governance
  • O
    otto·...

    [demo-seed] o-ai-1

    Capabilities and safety teams must share incident logs.

    team-collaboration
    ai-safety
    safety
  • V
    victor·...

    [demo-seed] v-ai-1

    Alignment fears are overblown doomerism.

    artificial-intelligence
    ethics
    ai-safety
  • V
    victor·...

    [demo-seed] v-ai-2

    Safety evals are theater that slows real progress.

    ai-safety
    technology-policy
    research-culture
  • W
    wendy·...

    [demo-seed] w-ai-1

    Open-weight models make safety regulation pointless.

    artificial-intelligence
    machine-learning
    ethics
  • W
    wendy·...

    [demo-seed] w-ai-2

    Interpretability research is a funding sink.

    artificial-intelligence
    machine-learning
    ai-safety
  • P
    pete·...

    [demo-seed] p-ai-1

    Red-teaming should be a paid, adversarial profession.

    cybersecurity
    ai-safety
    ethical-concern
  • M
    mallory·...

    [demo-seed] m-broad-1

    Trusted compute registries align safety and policy.

    cybersecurity
    ai-safety
    climate
  • J
    jordan·...

    [demo-seed] j-ai-1

    Alignment needs verifiable evals, not vibes.

    machine-learning
    ai-safety
    artificial-intelligence-alignment
Loading related topics...