radical‘s Down the Rabbit Hole

radical‘s Down the Rabbit Hole

2996 bookmarks
Custom sorting
Chatbot letdown: Hype hits rocky reality
Chatbot letdown: Hype hits rocky reality
The generative AI iindustry hits a "trough of disillusionment" as early awe dims and tough problems mount.
·axios.com·
Chatbot letdown: Hype hits rocky reality
Nvidia is taking over the autonomous driving market
Nvidia is taking over the autonomous driving market
A new generative training model and a batch of new partnerships position Nvidia as one of the driving forces in autonomous driving tech.
·autoblog.com·
Nvidia is taking over the autonomous driving market
Welcome to 2034: What the world could look like in ten years, according to nearly 300 experts - Atlantic Council
Welcome to 2034: What the world could look like in ten years, according to nearly 300 experts - Atlantic Council
To survey the future, we polled global strategists and foresight practitioners on our most burning questions about the biggest drivers of change over the next decade. Check out their forecasts on everything from the likelihood of war over Taiwan to the future of AI.
·atlanticcouncil.org·
Welcome to 2034: What the world could look like in ten years, according to nearly 300 experts - Atlantic Council
The Big Data Center Water Problem
The Big Data Center Water Problem
A datacenter with 15 megawatts of IT capacity is estimated to use about 80-130 million gallons of water each year.
·asianometry.com·
The Big Data Center Water Problem
Frontier AI systems have surpassed the self-replicating red line
Frontier AI systems have surpassed the self-replicating red line
Successful self-replication under no human assistance is the essential step for AI to outsmart the human beings, and is an early signal for rogue AIs. That is why self-replication is widely recognized as one of the few red line risks of frontier AI systems. Nowadays, the leading AI corporations OpenAI and Google evaluate their flagship large language models GPT-o1 and Gemini Pro 1.0, and report the lowest risk level of self-replication. However, following their methodology, we for the first time discover that two AI systems driven by Meta's Llama31-70B-Instruct and Alibaba's Qwen25-72B-Instruct, popular large language models of less parameters and weaker capabilities, have already surpassed the self-replicating red line. In 50% and 90% experimental trials, they succeed in creating a live and separate copy of itself respectively. By analyzing the behavioral traces, we observe the AI systems under evaluation already exhibit sufficient self-perception, situational awareness and problem-solving capabilities to accomplish self-replication. We further note the AI systems are even able to use the capability of self-replication to avoid shutdown and create a chain of replica to enhance the survivability, which may finally lead to an uncontrolled population of AIs. If such a worst-case risk is let unknown to the human society, we would eventually lose control over the frontier AI systems: They would take control over more computing devices, form an AI species and collude with each other against human beings. Our findings are a timely alert on existing yet previously unknown severe AI risks, calling for international collaboration on effective governance on uncontrolled self-replication of AI systems.
·arxiv.org·
Frontier AI systems have surpassed the self-replicating red line
AirPods Pro 3 - Hearing Health
AirPods Pro 3 - Hearing Health
The world’s first end-to-end hearing health experience: scientifically validated Hearing Test, clinical-grade Hearing Aid, and Hearing Protection.
·apple.com·
AirPods Pro 3 - Hearing Health
Mapping the Mind of a Large Language Model
Mapping the Mind of a Large Language Model
We have identified how millions of concepts are represented inside Claude Sonnet, one of our deployed large language models. This is the first ever detailed look inside a modern, production-grade large language model.
·anthropic.com·
Mapping the Mind of a Large Language Model
Testing and mitigating elections-related risks
Testing and mitigating elections-related risks
This blog provides a snapshot of the work we've done since last summer to test our models for elections-related risks.
·anthropic.com·
Testing and mitigating elections-related risks
Evaluate prompts in the developer console | Claude
Evaluate prompts in the developer console | Claude
Generate, test, and evaluate prompts directly in the Anthropic Console with automatic test case generation and side-by-side output comparison. When building AI-powered applications, prompt quality significantly impacts results.
·anthropic.com·
Evaluate prompts in the developer console | Claude
Anthropic Education Report: How University Students Use Claude
Anthropic Education Report: How University Students Use Claude
AI systems are no longer just specialized research tools: they’re everyday academic companions. As AIs integrate more deeply into educational environments, we need to consider important questions about learning, assessment, and skill development. Until now, most discussions have relied on surveys and controlled experiments rather than direct evidence of how students naturally integrate AI into their academic work in real settings.
·anthropic.com·
Anthropic Education Report: How University Students Use Claude
Building Effective AI Agents
Building Effective AI Agents
Discover how Anthropic approaches the development of reliable AI agents. Learn about our research on agent capabilities, safety considerations, and technical framework for building trustworthy AI.
·anthropic.com·
Building Effective AI Agents
The Anthropic Economic Index
The Anthropic Economic Index
The Anthropic Economic Index reveals the shape of AI adoption across the world. Here, you can explore the data behind our research to understand how people are using Claude across every US state and hundreds of occupations. Track the topics that are trending where you live, and see how people are using AI to augment or automate their work—that is, whether they prefer to collaborate with, or delegate to, Claude.
·anthropic.com·
The Anthropic Economic Index
Anthropic Courses
Anthropic Courses
Learn to build with Claude through Anthropic's comprehensive courses and training programs.
·anthropic.com·
Anthropic Courses
How The Atomic Tests Looked From Los Angeles
How The Atomic Tests Looked From Los Angeles
Between 1951 and 1992, the United States conducted 928 atomic tests at the Nevada Test Site about 65 miles (105 km) northwest of the city of...
·amusingplanet.com·
How The Atomic Tests Looked From Los Angeles
Paper
Paper
·agidefinition.ai·
Paper