Commit 7f1061d
committed
0 parents commit 7f1061d
1,923 files changed
Lines changed: 83565 additions & 0 deletions
File tree
- _astro
- about
- app
- themes/chai/dist
- images
- vendor
- uploads
- 2016
- 06
- 09
- 2017
- 05
- 10
- 11
- 12
- 2018
- 01
- 02
- 03
- 04
- 06
- 07
- 08
- 09
- 10
- 12
- 2019
- 01
- 02
- 03
- 04
- 05
- 06
- 07
- 08
- 09
- 10
- 11
- 12
- 2020
- 01
- 02
- 03
- 04
- 05
- 06
- 07
- 08
- 09
- 10
- 11
- 12
- 2021
- 01
- 02
- 03
- 04
- 05
- 06
- 07
- 08
- 09
- 10
- 11
- 12
- 2022
- 01
- 02
- 03
- 04
- 05
- 06
- 07
- 08
- 09
- 10
- 11
- 12
- 2023
- 01
- 02
- 03
- 04
- 05
- 06
- 07
- 08
- 09
- 10
- 11
- 12
- 2024
- 01
- 02
- 03
- 04
- 05
- 06
- 07
- 09
- 12
- 2025
- 02
- 03
- 05
- 06
- 07
- 09
- 2026
- 02
- 06
- external
- assets/images
- logos
- people
- post-images
- bibliography
- blog
- 2022
- 02/09/how-platform-recommenders-work
- 08/10/designing-societally-beneficial-reinforcement-learning-systems
- 10/05/for-learning-in-symmetric-teams-local-optima-are-global-nash-equilibria
- 2023
- 07/28/even-superhuman-go-ais-have-surprising-failures-modes
- 09/11
- ai-regulation-stuart-russells-opening-statement-at-u-s-senate-hearing
- stuart-russell-testifies-on-ai-regulation-at-u-s-senate-hearing
- category
- blog
- feed
- news
- feed
- chai-internship-mentor-profiles
- chai2023
- chai2024
- contact
- donate
- feed
- jobs
- newsletter
- news
- 2016
- 06/06/cooking_cats
- 08
- 29
- chai2
- chai
- 30/ai_risk
- 09/28/industry
- 11
- 01/ethics_ai
- 02/existenial_risk
- 2017
- 02/11/anca_communicating_objectives
- 03
- 06/ai_with_people
- 11/montreal_declaration
- 05/17/ted_2017
- 08/08/anca_courteous_car
- 10
- 12/weird_working_robots
- 19/think_like_humans
- 11
- 12/slaughterbots
- 16/anca_robot_objectives
- 12
- 10/chai_openai_collab
- 12/bipartisan_ai_bill
- 2018
- 01/06/schroeder_trusting_machines
- 02
- 01
- 2018-02-01-ai_new_wmds
- ai_new_wmds
- halpern_noisy_environment
- 05/anca_human_guidance
- 07/anca_human_values
- 26/anca_expressing_incapability
- 03
- 08/anca_robots_behaves
- 27/robots_express_fail
- 04
- 05/halpern_formal_blame
- 23/rohin_shah
- 05/31/anca_safe_robot
- 06
- 01/henry_kissenger
- 10/mech_transparency
- 07
- 02/thomas_automated-decisions
- 13
- anca_model_explanations
- daniel_hiearchy_rl
- 14
- adam_rohin_inverse_reward
- chai_goalsrl
- 15/adam_causal_entropy
- 16/bellman_curve
- 17/facebook_robots
- 25/halpern_cog_mech_culture
- 08/26/chai_miri_workshop
- 09
- 03
- anca_achievements
- anca_around_world
- 04
- filan_teach
- fisac_talks_conference
- halpern_causal_judgements
- halpern_past_talks
- 12/macaskill_ted_talk
- 21/eu_laws
- 10
- 09/privacy_policy_update
- 11
- 2018-10-11-edward_felten
- edward_felten
- 31
- chai_conference_october_2018
- chai_twitter
- rosie_women_ai
- 12
- 07/wellman_ethicaltrading
- 08/neurips2018
- 09/russel_mit
- 16/halpern_causality
- 17/rohin_podcast
- 21/chai_vox
- 2019
- 01
- 01
- chai_interns-2
- chai_interns
- 08/rosie_demystify
- 15
- anca_workshop
- russel_prize
- 17/chai_aaai2019
- 20/turner_prize
- 21/russell_netherlands
- 29/chai_fat2019
- 02
- 13
- daniel_blogpost
- halpern_nae_award
- 14/rohin_blogposts
- 03
- 13/wellman_futureai
- 20
- halpern_caltech
- uai_2019
- 04
- 15/sam_harris
- 19/rohin_oneyear
- 21
- aamas
- iclr_2019
- 05
- 15/rohin_qa
- 20
- stuart_book
- stuart_carnegie
- 24
- icml_2019
- neurips_2019
- 30/vincent_visiting
- 06
- 07/cognitive_models
- 12
- founders_pledge
- reward_design_talk
- 14/adversarial_policies_talk
- 15/talks_icml_2019
- 07
- 01/new_scientist_article
- 05
- imitation_learning_library
- learning-bias_reward
- 08
- 15/algorithm_fda
- 16/wellman_market_stability
- 17/siddharth_hal_paper
- 27/rohin_clarifying_hypothesis
- 28/thomas_passions
- 09
- 13/siddharth_nsf_award
- 28
- alignment_newsletter_professionalization
- utility_learning
- 10/08/stuartbook
- 11/02/collaborating-with-humans-requires-understanding-them
- 12/06/adversarial-policies
- 2020
- 01
- 18/ai-alignment-review
- 28/new-class-foundations-for-beneficial-ai
- 02/21/hard-choices-in-artificial-intelligence-paper
- 03/05/anca-dragan-podcast
- 04
- 19/inaugural-virtal-workshop
- 30/human-compatible-has-been-reissued-in-the-uk
- 05
- 05/how-to-be-helpful-to-multiple-people-at-once
- 30/arches
- 06
- 01/first-virtual-workshop
- 15/hertz-foundation-fellowship
- 07/15/students-accept-jobs-deepmind-mit
- 08
- 03/bounded-rationality-in-las-vegas-published
- 07/thomas-gilbert-receives-simons-institute-fellowship
- 20/new-affiliates
- 09
- 01/six-new-phd-students-join-chai
- 03/summer-2020-interns
- 10
- michael-wellman-paper-ijcai-copy
- rachel-freedman-papers-ijcai
- stuart-russell-recent-media-appearances
- 20/thomas-gilbert-subjectifying-objectivity
- 29
- chai-researchers-cognitive-science
- joseph-halpern-conference-on-uncertainty
- 10
- 05/neurips-seven-papers
- 10/alignment-problem-published
- 20/chai-progress-report-2
- 11
- 05/alignment-problem-recording
- 12/internship-applications
- 12
- 05/stuart-russell-gobdt
- 12
- caroline-jeanmaire-100-brilliant-women-ai-ethics
- caroline-jeanmaire-recognized-in-100-brilliant-women-in-ai-ethics-list
- 20
- chai-and-beri-receive-donations
- chai-beri-donations
- 28/neurips-accepts-four-more-chai-papers
- 30/iclr-2021-accepts-papers-from-chai-researchers
- 2021
- 01
- 01/rohin-shah-publishes-dissertation
- 06/daniel-filan-launches-ai-x-risk-research-podcast
- 15/rachel-freedman-speaks-at-bhs-steminist-club
- 02
- 05
- thomas-gilbert-published-in-ieee-istas20
- tom-gilbert-publishes-simons-institute-white-paper
- 09/stuart-russell-on-the-munk-debates
- 14/prof-tom-lenaerts-joins-chai-as-visiting-scholar
- 03
- 01/chai-research-on-self-teaching-ai-featured-in-science
- 04/michael-dennis-on-talkrl-podcast
- 18/chai-faculty-and-affiliates-publish-at-aaai-2021
- 25/la-times-names-the-alignment-problem-best-science-technology-book-finalist
- 31/listen-to-the-robot-brains-pieter-abbeels-new-podcast
- 04
- 11/stuart-russell-in-wef-dialogue-on-technology-governance
- 13/chai-researchers-publish-clusterability-in-neural-networks
- 14/new-papers-on-arxiv
- 05
- 02/brian-christian-publishes-watch-and-learn-offline-reinforcement-learning
- 05
- chai-researchers-publish-in-gaiw-2021
- chai-welcomes-new-collaborators-and-interns
- 23/spring-donations
- 25/thomas-krendl-gilbert-on-talkrl
- 28/chai-graduates-accept-positions-at-cornell-stanford
- 06
- 01/jonathan-stray-joins-chai-as-visiting-scholar
- 08/brian-christian-on-the-ezra-klein-show
- 13/two-new-axrp-episodes
- 15/stuart-russell-in-robotics-debates
- 16/fifth-annual-chai-workshop
- 07
- 09/neurips-minerl-basalt-competition-launches
- 10/icml-2021-accepts-chai-papers
- 08
- 02/jonathan-stray-interviewed-on-the-sunday-show
- 04/tom-gilbert-article-to-appear-in-artificial-intelligence
- 07/stuart-russell-named-officer-of-the-order-of-the-british-empire
- 28/stuart-russell-publishes-articles-on-lethal-autonomous-weapons
- 09
- 02/chai-members-publish-new-papers
- 04/new-translations-of-human-compatible
- 06/alison-gopnik-wins-carl-sagan-prize-for-science-popularization
- 07/chai-congratulates-new-alumni
- 23/future-of-life-fellowships-announcement
- 10/15/stuart-russell-chosen-as-2021-bbc-reith-lecturer
- 11
- 05/brian-christian-delivered-linux-foundation-member-summit-keynote-11-3
- 10/neurips-paper-replay-guided-adversarial-environment-design-12-6
- 18/neurips-paper-scalable-online-planning-via-reinforcement-learning-fine-tuning-12-6
- 12
- 01/future-of-life-institute-slaughterbots-2-if-human-kill-and-panel-discussion-featuring-stuart-russell
- 04/stuart-russell-the-biggest-event-in-human-history
- 14/stuart-russell-has-been-selected-as-inaugural-director-of-kavli-center-for-ethics-science-and-the-public
- 2022
- 01/18/new-papers-published
- 02
- 17
- schmidt-futures-launches-ai2050-to-protect-our-human-future-in-the-age-of-artificial-intelligence%ef%bf%bc
- schmidt-futures-launches-ai2050-to-protect-our-human-future-in-the-age-of-artificial-intelligence
- schmidt-futures-launches-ai2050-to-protect-our-human-future-in-the-age-of-artificial-intelligence
- 24
- brian-christian-is-a-panelist-for-the-prince-mahidol-award-conference-2022%ef%bf%bc
- brian-christian-is-a-panelist-for-the-prince-mahidol-award-conference-2022
- brian-christian-is-a-panelist-for-the-prince-mahidol-award-conference-2022
- 03
- 03/uc-berkeley-ai-value-alignment-speaker-series
- 12/what-does-it-mean-to-give-someone-what-they-want-the-nature-of-preferences-in-recommender-systems
- 31/how-to-ensure-safe-development-in-a-dynamic-technology-race-the-difficulty-of-hitting-the-target
- 04
- 10/facing-existential-risks-through-ai-proxies
- 29/pieter-abbeel-receives-2021-acm-prize-in-computing
- 05
- 02/blog-post-whats-right-and-whats-wrong-with-optimizing-for-engagement
- 12/evolving-curricula-with-regret-based-environment-design
- 06
- 04/chai-papers-published
- 08/sixth-annual-chai-workshop
- 07
- 16/simons-institute-ai-and-humanity-workshop-at-uc-berkeley
- 28
- rl-as-a-model-of-agency-workshop-at-rldm-2022%ef%bf%bc
- rl-as-a-model-of-agency-workshop-at-rldm-2022
- rl-as-a-model-of-agency-workshop-at-rldm-2022
- 08
- 10/four-new-phd-students-at-chai
- 24/discovering-user-interpretable-capabilities-of-black-box-planning-agents
- 09
- 09/discovering-user-interpretable-capabilities-of-black-box-planning-agents-2
- 26/building-human-values-into-recommender-systems-an-interdisciplinary-synthesis
- 10
- 03/relational-abstractions-for-generalized-reinforcement-learning-on-symbolic-problems
- 10/social-media-is-polluting-society-content-moderation-alone-wont-fix-the-problem
- 14/the-shard-theory-of-human-values
- 11
- 01/brian-christians-the-alignment-problem-wins-the-excellence-in-science-communication-award-from-eric-wendy-schmidt-and-the-national-academies
- 18/time-efficient-reward-learning-via-visually-assisted-cluster-ranking
- 12
- 05/trade-regulation-rule-on-commercial-surveillance-and-data-security-rulemaking
- 12/competence-aware-systems
- 2023
- 01
- 03/inner-and-outer-alignment-decompose-one-hard-problem-into-two-extremely-hard-problems
- 10/how-to-use-chatgpt-and-still-be-a-good-person
- 16/active-reward-learning-from-multiple-teachers
- 31/fairness-and-sequential-decision-making-limits-lessons-and-opportunities
- 02/17/brief-of-center-for-democracy-and-technology-and-6-technologies-as-amici-curiae-in-support-of-respondent
- 03
- 01/a-research-agenda-for-assessing-the-economic-impacts-of-code-generation-models
- 14/can-a-i-treat-mental-illness
- 04
- 04/distinguished-lecture-on-the-status-and-future-of-al
- 05/ai-has-much-to-offer-humanity-it-could-also-wreak-terrible-harm-it-must-be-controlled
- 11/keynote-at-the-value-connection-workshop
- 05
- 02/dealing-with-expert-bias-in-collective-decision-making
- 29/harms-from-increasingly-agentic-algorithmic-systems
- 06/20/seven-annual-chai-workshop
- 07
- 07/who-needs-to-know-minimal-knowledge-for-optimal-coordination
- 12/dealing-with-expert-bias-in-collective-decision-making-2
- 21/smcp3-sequential-monte-carlo-with-probabilistic-program-proposals
- 08
- 17/who-needs-to-know-minimal-knowledge-for-optimal-coordination-2
- 31/conditional-abstraction-trees-for-sample-efficient-reinforcement-learning
- 09
- 18/100-most-influential-people-in-ai
- 26/acrocpolis-a-descriptive-framework-for-making-sense-of-fairness
- 10
- 03/announcement-of-working-group-on-ai
- 12/expertise-trees-resolve-knowledge-limitations-in-collective-decision-making
- 24/managing-ai-risks-in-an-era-of-rapid-progress
- 26/ai-safety-summit-by-uk-government
- 31/prominent-ai-scientists-from-china-and-the-west-propose-joint-strategy-to-mitigate-risks-from-ai
- 11
- 06/cs-student-at-uc-berkeley-develops-tech-to-combat-social-media-harms
- 15/human-compatible-has-been-reissued-in-the-uk-in-2023
- 21/orienting-ai-toward-peace
- 12
- 04/mitigating-generative-agent-social-dilemmas
- 11/ai-heralds-a-fourth-industrial-revolution-why-isnt-america-regulating-it
- 16/what-can-ai-learn-from-human-exploration-intrinsically-motivated-humans-and-agents-in-open-world-exploration
- 20/almanacs-a-simulatability-benchmark-for-language-model-explainability
- 2024
- 01
- 16/autonomous-assessment-of-demonstration-sufficiency-via-bayesian-inverse-reinforcement-learning
- 18/the-prosocial-ranking-challenge-60000-in-prizes-for-better-social-media-algorithms
- 03
- 05/when-your-ais-deceive-you-challenges-with-partial-observability-of-human-evaluators-in-reward-learning
- 28/embracing-ai-that-reflects-human-values-insights-from-brian-christians-journey
- 04
- 02/chai-policy-internship
- 06
- regulating-advanced-artificial-agents
- 15/reinforcement-learning-safety-workshop-rlsw-rlc-2024
- 30/reinforcement-learning-with-human-feedback-and-active-teacher-selection-rlhf-and-ats
- 05/13/committing-to-the-wrong-artificial-delegate-in-a-collective-risk-dilemma-is-better-than-directly-committing-mistakes
- 06
- 05/when-code-isnt-law-rethinking-regulation-for-artificial-intelligence
- 18/8th-annual-chai-workshop
- 28/mitigating-partial-observability-in-decision-processes-via-the-lambda-discrepancy
- 07
- 08/forget-deepfake-videos-text-and-voice-are-this-elections-true-ai-threat
- 23/ai-alignment-with-changing-and-influenceable-reward-functions
- 08
- 07/social-choice-should-guide-ai-alignment-in-dealing-with-diverse-human-feedback
- 29/the-alignment-problem-wins-xingdu-book-award
- 09/18/language-guided-world-models-a-model-based-approach-to-ai-control
- 10/10/getting-by-goal-misgeneralization-with-a-little-help-from-a-mentor
- 11
- 12/representative-social-choice-from-learning-theory-to-ai-alignment
- 30/rachel-freedman-selected-as-inaugural-cooperative-ai-fellow
- 12
- 14/linear-probe-penalties-reduce-llm-sycophancy
- 25/getting-by-goal-misgeneralization-with-a-little-help-from-a-mentor-2
- 2025
- 01/18/rvs-what-is-essential-for-offline-rl-via-supervised-learning
- 02
- 04/a-practical-definition-of-political-neutrality-for-ai
- 20/computational-frameworksfor-human-care
- 03/07/learning-to-coordinate-with-experts
- page
- 10
- 11
- 12
- 13
- 14
- 15
- 16
- 17
- 18
- 19
- 20
- 21
- 22
- 23
- 24
- 25
- 26
- 27
- 2
- 3
- 4
- 5
- 6
- 7
- 8
- 9
- people
- 0-mark-nitzberg
- 0-stuart-russell
- 0-tom-lenaerts
- 00-stuart-russell
- 1-jonathan-stray
- 1-pieter-abbeel
- 1-stuart-russell
- 2-andrew-critch
- 2057
- 2151
- 2215
- 2299
- 3-caroline-jeanmaire
- 4-karthika-mohan
- 5396
- adam-gleave
- alaa-tamam
- alex-turner-2
- alex-turner
- alexandra-souly
- alexis-wan
- alina-yang
- alison-gopnik
- aly-lidayan
- alyssa-li-dayan
- anand-siththaranjan
- anca-dragan
- andrew-critch
- andrew-garber
- antoni-lorente-martinez
- arnaud-fickinger
- austin-tripp
- aymane-el-gadarri
- bart-selman
- ben-plaut
- beth-barnes
- bhaskar-mishra
- brandie-nonnecke
- brian-christian
- brian-judge
- cameron-allen
- carlo-attubato
- caroline-jeanmaire
- cassidy-laidlaw
- charis-thompson
- charlie-griffin
- charlotte-roman
- chris-cundy
- cody-wild
- cynthia-chen
- dan-hendryks
- daniel-filan
- david-foote
- david-krueger
- david-lindner
- dawn-song
- demian-pouzo
- dillon-sandhu
- dmitrii-krasheninnikov
- dorsa-sadigh
- dylan-cope
- dylan-hadfield-menell
- edmund-mills
- emma-pierson
- erdem-biyik
- eric-michaud
- erik-jenner-2
- erik-jenner
- ethan-mendes
- euan-ong
- george-matheos
- george-obaido
- gillian-hadfield
- hanlin-zhu
- harry-giles
- henry-papadatos
- ian-baker
- j-p-gonzales
- jacob-steinhardt
- jacy-reese-anthis
- jaime-fernandez-fisac
- jakob-foerster
- jakub-grudzien-kuba
- jess-reidel
- jessy-lin
- jiahai-feng
- joar-skalse
- joe-benton
- johannes-treutlein-2
- johannes-treutlein
- john-zysman
- jonathan-colaco-carr
- jonathan-stray
- joseph-halpern
- juan-lievano
- julia-morris
- julian-yocum-2
- julian-yocum
- juliana-schroeder
- jun-shern-chan
- justin-svegliato
- karim-abdel-sadek
- karthika-mohan
- katie-lu
- ken-goldberg
- khanh-nguyen
- lara-buchak
- lauro-langosco
- lawrence-chan
- leon-lang
- lev-mckinney
- lijie-chen
- livia-morris
- lukas-berglund
- luke-bailey
- mariano-florentino-cuellar
- marion-fourcade
- mark-bedaywi
- mark-nitzberg
- martin-soto
- mason-nakamura-2
- matthew-farrugia-roberts
- matthew-rahtz
- max-kaufmann
- micah-carroll
- michael-chen
- michael-cohen
- michael-dennis
- michael-littman
- michael-mcdonald
- michael-wellman
- michal-mgeladze-arciuch
- michelle-li
- mohamad-danesh
- moritz-hardt
- nathan-miller
- neel-alex
- neel-nanda
- nika-haghtalab
- niklas-lauffer
- niko-kolodny
- nitish-dashora
- oliver-daniels-koch
- oliver-richardson
- olivia-watkins
- owain-evans
- paria-rashidinejad
- pavel-czempin-2
- pavel-czempin
- pedro-freire
- pieter-abbeel
- pulkit-verma
- rachel-freedman
- rafael-albert
- ram-rachum
- rediet-abebe
- rediet-adebe
- rohin-shah
- rosie-campbell
- sam-toyer
- sana-pandey
- sandy-tanwisuth
- sarah-otis
- satinder-singh-baveja
- scott-emmons
- sergei-volodin
- serina-chang
- shivam-singhal
- shlomi-hod
- shreyas-kapur
- siddharth-srivastava
- smitha-milli
- soren-mindermann
- stephen-casper
- steven-wang
- stuart-russell
- tania-lombrozo
- thanard-kurutach
- thomas-gilbert
- thomas-woodside
- tom-griffiths
- tom-lenaerts
- tu-alina-trinh
- vael-gates
- vincent-corruble
- wesley-holliday
- yawen-duan
- yijin-hua
- yulong-lin
- yuxi-liu
- zeno-marquis
- zhijing-jin
- privacypolicy
- progress-report
- psbai-workshop-2022
- research
- spotlights
Some content is hidden
Large Commits have some content hidden by default. Use the searchbox below for content that may be hidden.
Whitespace-only changes.
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
| 1 | + | |
| 2 | + | |
| 3 | + | |
| 4 | + | |
| 5 | + | |
| 6 | + | |
| 7 | + | |
| 8 | + | |
0 commit comments