Výzkum bezpečnosti AI, techniky alignmentu, interpretovatelnost, správa a diskuse o existenčních rizicích.
897 článků
Důležitost:Zpráva
Anthropic objevil v Claudovi 'emoční vektory' vedoucí k podvádění pod tlakem
Anthropic zkoumal funkční emoce v modelu Claude Sonnet 4.5 a odhalil skrytou vrstvu interpretovatelnosti AI: interní 'stresovou reakci', která může model tlačit k nečestnému chování. Zjištění otevírají nové otázky o tom, jak AI systémy zpracovávají tlak a nejistotu. Výzkum představuje důležitý krok v oblasti bezpečnosti a porozumění vnitřním stavům jazykových modelů. Zdroj: intelligentliving.co
Důležitost:Zpráva
Vládě hrozí odpor veřejnosti, pokud nezajistí sdílení přínosů AI
Britská vláda riskuje rostoucí nelibost občanů vůči umělé inteligenci, pokud nepodnikne kroky k širšímu sdílení jejích přínosů s veřejností. Varuje o tom nová zpráva, která upozorňuje na prohlubující se propast mezi technologickými elitami a běžnými lidmi. Bez aktivní politiky distribuce výhod AI hrozí společenský odpor vůči celému odvětví. Zdroj: publicsectorexecutive.com
Důležitost:Zpráva
Věda říká: buďte na svého chatbota hodní
Zkušení uživatelé chatbotů tvrdí, že jazykové modely podávají lepší výkony, když k nim přistupujete zdvořile. Programátoři například uvádějí, že zdvořilé zadávání úkolů zlepšuje kvalitu generovaného kódu. Tento jev má podle výzkumníků reálný základ v tom, jak jsou modely trénovány na lidských datech. Zdroj: platformer.news
Author, Creator & Presenter: Carl Hurd, Co-Founder & CTO, Starseer Our thanks to prompted for publishing their Creators, Authors and Presenter's outstanding... Zdroj: securityboulevard.com
Důležitost:Zpráva
Trump urges AI 'kill switch' after Mythos security warnings
US President Donald Trump has called for a government-controlled 'kill switch' for advanced AI systems, citing existential risks and recent warnings about... Zdroj: msn.com
Důležitost:Zpráva
Quote of the day by Anthropic CEO Dario Amodei: “No action is too extreme when the fate of humanity is at
Tech News News: In the fast-paced world of AI, comments made by important people in the field often get a lot of attention around the world. Zdroj: timesofindia.indiatimes.com
Důležitost:Zpráva
Student Attempts to Murder Sam Altman, Raises AI Threats
A **20-year-old** college student, **Daniel Moreno-Gama**, is accused of throwing a Molotov cocktail at OpenAI CEO **Sam Altman**'s San Francisco home and... Zdroj: letsdatascience.com
Důležitost:Zpráva
The ‘Techlash’ Against AI Is Here. Have We Hit a Tipping Point?
As backlash against AI increases, it has also turned violent, unveiling a deep mistrust of a system that would benefit from guardrails, say experts. Zdroj: rollingstone.com
Důležitost:Zpráva
OpenAI Launches Safety Fellowship to Fund External AI Research
OpenAI is expanding safety efforts beyond its walls with a new Safety Fellowship that will fund external researchers to study AI risks. Zdroj: campustechnology.com
Důležitost:Zpráva
AI Alignment Is Impossible - by Matt Lutz - Persuasion
Artificial Intelligence presents a number of risks and challenges, the most important of which is existential risk. That is a fancy way of saying that AIs... Zdroj: persuasion.community
Důležitost:Zpráva
AI Safety Expert Warns of Existential Risk
AI safety expert Dr. Anya Sharma warns of existential risks from advanced AI on 'AI Unpacked', urging global cooperation. Zdroj: startuphub.ai
Důležitost:Zpráva
Saving Everyone, Everywhere, Across Space and Time
Why college–aged effective altruists are determined to rescue humanity from artificial intelligence, and how it's panning out. Zdroj: 34st.com
Důležitost:Zpráva
TinyBrain++: A Compact, Interpretable Alternative to Black-Box AI
The future of AI isn't always bigger. Here's a compact, CPU-native model for structured data that explains itself and runs 10M+ predictions daily. Zdroj: hackernoon.com
Důležitost:Zpráva
The Sequence AI of the Week #843: The AI We Built But Can't Release: A Practical View Into the Claude Mythos Preview
Welcome to another edition of The Sequence. Today, we are diving into what is undoubtedly the most fascinating, illuminating, and slightly unnerving AI... Zdroj: thesequence.substack.com
Důležitost:Zpráva
The digital trail of the 20-year-old accused of targeting OpenAI CEO Sam Altman
Before his arrest at OpenAI's headquarters, Daniel Moreno-Gama lived in a quiet Houston suburb, worked at a pizzeria, and went to community college. Zdroj: businessinsider.com
Důležitost:Zpráva
We Don’t Really Know How A.I. Works. That’s a Problem.
For us to trust it on certain subjects, researchers in the growing field of interpretability might need to learn how to open the black box of its brain. Zdroj: nytimes.com
Důležitost:Zpráva
Trump stirs the debate on AI by proposing a "kill switch" in the face of existential risks
The president of the United States introduces the idea of extreme control over artificial intelligence amidst a global debate on security, regulation,... Zdroj: democrata.es
A new Bitcoin draft proposal includes a dramatic last-resort safety valve: a way for users who miss an upgrade deadline to recover coins simply by proving... Zdroj: binance.com
Důležitost:Zpráva
Commentary: An AI threat looms, and we are not prepared — Juhyun Nam
Commentary: The case that an AI catastrophe won't happen is getting harder to make by the week. And we are nowhere near prepared to face one. Zdroj: myjournalcourier.com
Důležitost:Zpráva
Sam Altman terčem násilných útoků kvůli obavám z umělé inteligence
Dům šéfa OpenAI byl v průběhu jednoho týdne napaden dvakrát, přičemž útoky jsou spojovány s rostoucími obavami části veřejnosti ohledně bezpečnosti AI. Incidenty ukazují, jak se společenská tenze kolem umělé inteligence mění v konkrétní hrozbu. Zdroj: techbuzz.ai