lesswrong.com
75 articles in the Truth Foundry index. Each links out to the original.
- Ethical Consumption Needs to Be Affordable Amid Economic Downturn · Ethical consumption must become accessible and affordable as the economy worsens.
- LessWrong Critiques Flaws in Current AI Safety Bug Bounty Programs · LessWrong article argues existing AI safety bug bounties fail to incentivize meaningful research and expose critical vulnerabilities.
- Researchers Monitor Computer-Use Agents in OSWorld-Control Environment · LessWrong article discusses monitoring computer-use agents within the OSWorld-control simulation environment.
- Alignment to What? Examining the Core Question of AI Safety · The article questions the fundamental assumption of AI alignment by asking what specific goals AI should be aligned to.
- LessWrong Debate: Popperians, Bayesians, and Ramseyians on Rationality · LessWrong community discusses epistemological frameworks including Popperian falsification, Bayesian updating, and Ramseyian decision theory.
- Prefixing Names with 'secure_' Improves AI Code Security · Using 'secure_' prefixes in variable names helps AI agents generate more secure code.
- Can Large Language Models Effectively Teach Students? · The article explores the feasibility and limitations of using LLMs as educational tools.
- Systems Dynamics Model for Pausing AI Development · A systems dynamics framework proposes pausing AI development to prevent runaway feedback loops.
- AI Companions May Harm Social Skills Rather Than Improve Them · Interacting with AI companions could degrade human social skills instead of enhancing them.
- Contagious Humming Technique Silences Rooms Instantly · A viral video demonstrates a humming technique that rapidly quiets a noisy room.
- LessWrong Article Discusses Dissolving Deep Learning Sample Efficiency Gap · A LessWrong post explores theoretical approaches to closing the sample efficiency gap in deep learning.
- LessWrong Advocates for Breadth-First AI Safety Planning Approaches · LessWrong argues for breadth-first AI safety plans to address diverse risks simultaneously.
- Uploaded Minds Would Self-Download in xP-Zombie Scenarios · Uploaded consciousnesses could theoretically transfer themselves to new bodies without external intervention.
- Senator Sanders Proposes Government Take 50% Stake in AI Labs · Senator Bernie Sanders proposes a bill requiring the U.S. government to hold a 50% ownership stake in major AI research laboratories.
- Opus 4.8 Part 2 Examines Model Welfare and Alignment Risks · The article analyzes the concept of model welfare in advanced AI systems.
- AIGS Canada's Remarkable Story Explored on LessWrong · An article on LessWrong details the unique history and development of AIGS Canada.
- Superintelligence of the gaps: A LessWrong analysis of AI limitations · The article explores the concept of superintelligence emerging from gaps in current AI capabilities.
- Lean, not backpressure: A strategy for managing cognitive load · The article argues that reducing cognitive load through 'leaning' is more effective than relying on backpressure mechanisms.
- Humans Can Be Hermaphroditic and Self-Fertilize, Though It Shouldn't Happen · The article explains that some humans are hermaphroditic and can self-fertilize, but this is rare and generally discouraged.
- LessWrong Author Reflects on Underestimating AI Capabilities · A LessWrong contributor shares personal reactions to the recurring theme of underestimating artificial intelligence.
- Understanding Survey Data: Why Lizardmen Are Not Constant · The article explains that survey data distributions shift over time, challenging the assumption of static populations.
- Why the 'Unrealistic Hypothetical' Objection Fails in Rational Discourse · The author argues that dismissing hypothetical scenarios as unrealistic is a logical fallacy that hinders effective reasoning.
- LessWrong introduces xNLA Thought Anchors for reasoning stability · LessWrong launches xNLA Thought Anchors to stabilize reasoning in complex AI models.
- Feasibility Study for Lighthaven East Naval Base Expansion · A feasibility study evaluates the strategic viability of expanding the Lighthaven East naval facility.
- LessWrong Explores Barriers to a Prosperous Future · LessWrong discusses obstacles preventing a prosperous future for humanity.
- Analysis of Axes of Variation in Third-Party Risk Assessment · The article explores dimensions for evaluating risks posed by third-party entities.
- Automated AI Production May Lead to Concentration of Power · Automated AI production could concentrate power in the hands of a few dominant entities.
- Researchers Visualize Cyclical Structure in Llama Model Architecture · Analysis reveals hidden cyclical patterns within the transformer architecture of the Llama language model.
- Author Argues AI Evaluation Metrics Are Most Valuable Research Focus · The author contends that improving AI evaluation benchmarks is the most critical area for current AI research efforts.
- LessWrong Explores Atmospheric Water Generation for Desert Sustainability · LessWrong discusses extracting food, water, and power from thin desert air.