Zotero Catalogue

A public reading list of papers, books, videos, and other resources. The inclusion of a resource on this catalogue is NOT an endorsement of anything contained within, and in most cases the resources has not been read by me at the time of saving.

1157 items · showing 401–450 · page 9 of 24 Sort: Newest Oldest Title A–Z Title Z–A

ZIMV72IY
book
For a new liberty: the libertarian manifesto
Murray N. Rothbard, Llewellyn H., Jr Rockwell
2006 · Ludwig von Mises Inst.
Saved 2026-07-27
22NAHZ4D
book
The myth of the rational voter: why democracies choose bad policies
Bryan Douglas Caplan
2008 · Princeton University Press
Saved 2026-07-27
"Caplan argues that voters continually elect politicians who either share their biases or else pretend to, resulting in bad policies winning again and again by popular demand. Calling into question our most basic assumptions about American politics, Caplan contends that democracy fails precisely because it does what voters want. Through an analysis of American's voting behavior and opinions on a range of economic issues, he makes the case that noneconomists suffer from four prevailing biases: they underestimate the wisdom of the market mechanism, distrust foreigners, undervalue the benefits of conserving labor, and pessimistically believe the economy is going from bad to worse. Caplan lays out several ways to make democratic government work better
PL3VX7YH
book
The socialist manifesto the case for radical politics in an era of extreme inequality
Bhaskar Sunkara
2020 · Verso
Saved 2026-07-27
8QKQ88IR
book
The long depression: how it happened, why it happened, and what happens next
Michael Roberts
2016 · Hyamarket Books
Saved 2026-07-27
9KFQCRD2
book
Capital
Karl Marx
Saved 2026-07-27
EKWDW2YM
webpage
Dejan Panovski
2018
Saved 2026-07-27
SCP copies files securely between local and remote hosts over SSH. This guide covers syntax, common options, and practical examples for everyday file transfers.
MTP7BISW
journalArticle
But Who will Monitor the Monitor?
David Rahman
Saved 2026-07-26
Consider a group of individuals in a strategic environment with moral hazard and adverse selection, and suppose that providing incentives for a given outcome requires a monitor to detect deviations. What about the monitor’s deviations? In this paper I propose a contract that makes the monitor responsible for the monitoring technology, and thereby successfully provides incentives even when the monitor’s observations are not only private, but costly, too. I also characterize exactly when such a contract can provide monitors with the right incentives to perform. In doing so, I emphasize virtual enforcement and suggest its implications for the theory of repeated games.
A8D2A4C4
journalArticle
David Rahman
2012 · American Economic Review
Saved 2026-07-26
Suppose that providing incentives for a group of individuals in a strategic context requires a monitor to detect their deviations. What about the monitor's deviations? To address this question, I propose a contract that makes the monitor responsible for monitoring, and thereby provides incentives even when the monitor's observations are not only private, but costly, too. I also characterize exactly when such a contract can provide monitors with the right incentives to perform. In doing so, I emphasize virtual enforcement and suggest its implications for the theory of repeated games. (JEL C78, D23, D82, D86)
6YYJMJVG
webpage
Saved 2026-07-26
steerable and explainable AI
VYNJUDYB
webpage
Lila Shroff, Rose Horowitch
2026
Saved 2026-07-26
AI companies are stripping universities of their best researchers.
UF2KXW3N
preprint
Jiarui Zhang, Muzi Tao, Shangshang Wang, Ollie Liu, Xuezhe Ma, Willie Neiswanger
2026
Saved 2026-07-24
Human vision is a closed loop: gaze is continuously redirected by intermediate hypotheses rather than a single snapshot. Decades of psychophysics and cognitive science have argued that this active observation is essential for a wide range of tasks. Whether today's multimodal large language models (MLLMs) exercise active observation is an empirical question that current vision-language benchmarks do not answer. We introduce ActiveVision, a benchmark that makes active observation measurable for MLLMs, comprising 17 tasks across 3 categories. Tasks are designed to force repeated visual perception rather than a single static description. Frontier MLLMs collapse on ActiveVision: the highest-scoring model we evaluate, GPT-5.5 at the highest exposed reasoning-effort tier, solves only 10.6% of items and scores zero on 11 of the 17 tasks, and even Claude Fable 5, despite topping most reasoning and coding leaderboards, solves just 3.5%, far behind three human participants who average 96.1%. Furthermore, much of the gap persists even when models write and run their own vision code: such code is unreliable on realistic imagery, and catching its failures itself requires the active perception the models lack. Together, these results indicate that current MLLMs lack robust active visual observation, motivating architectures and training objectives that close the perception-reasoning loop.
KW8BXXF5
forumPost
Felix Choussat
2026
Saved 2026-07-23
C8EEHYDF
journalArticle
International AI Safety Report 2026
2026
Saved 2026-07-23
ENLRWTQU
blogPost
2023
Saved 2026-07-23
Why are billions of dollars being poured into artificial intelligence R&D this year? Companies certainly expect to get a return on their investment. Arguably, the main reason AI is profitable i…
687X8XDT
webpage
Saved 2026-07-23
SECDNSGM
preprint
Joseph Carlsmith
2024
Saved 2026-07-23
This report examines what I see as the core argument for concern about existential risk from misaligned artificial intelligence. I proceed in two stages. First, I lay out a backdrop picture that informs such concern. On this picture, intelligent agency is an extremely powerful force, and creating agents much more intelligent than us is playing with fire -- especially given that if their objectives are problematic, such agents would plausibly have instrumental incentives to seek power over humans. Second, I formulate and evaluate a more specific six-premise argument that creating agents of this kind will lead to existential catastrophe by 2070. On this argument, by 2070: (1) it will become possible and financially feasible to build relevantly powerful and agentic AI systems; (2) there will be strong incentives to do so; (3) it will be much harder to build aligned (and relevantly powerful/agentic) AI systems than to build misaligned (and relevantly powerful/agentic) AI systems that are still superficially attractive to deploy; (4) some such misaligned systems will seek power over humans in high-impact ways; (5) this problem will scale to the full disempowerment of humanity; and (6) such disempowerment will constitute an existential catastrophe. I assign rough subjective credences to the premises in this argument, and I end up with an overall estimate of ~5% that an existential catastrophe of this kind will occur by 2070. (May 2022 update: since making this report public in April 2021, my estimate here has gone up, and is now at >10%.)
8V4UTSEL
webpage
Saved 2026-07-23
New research on how we've reduced agentic misalignment
DUZQSQUI
webpage
2026
Saved 2026-07-23
IUYNJRGC
journalArticle
New Generation of Counter UAS Systems to Defeat of Low Slow and Small (LSS) Air Threats
Jacco Dominicus
Saved 2026-07-22
Detecting, classifying, identifying, tracking and defeating low, slow and small air threats presents a major challenge for existing sensor and effector systems. So-called first generation Counter Unmanned Aircraft Systems (C-UAS) systems often rely on detecting the datalink from the controller to the drone which provides limited capability against current threats. However, this means of detecting drones is a challenge when operators manipulate standard datalinks and it will not work at all against current and future autonomous drones. Other current methods of detecting and neutralising drones include for example combining radar with optical sensors. These systems are not always reliable, can generate large numbers of false alerts and are often manpower intensive to operate. The NATO SCI-301 Research Task Group (RTG) has been working on specifying what second generation C-UAS systems should entail. This paper will outline the findings of this RTG over the past three years.
MPAC8FIT
forumPost
Scott Alexander
2009
Saved 2026-07-22
JF4F4CZE
forumPost
Yair Halberstadt
2026
Saved 2026-07-22
UDPJBLKL
blogPost
Aella
2025
Saved 2026-07-22
The Growing Kids God's Way protocol
K7WRRPIW
blogPost
Aella
2025
Saved 2026-07-22
and the way we treat children as property
8QQZKPHP
blogPost
Saved 2026-07-22
DUV3PX99
journalArticle
A REPORT TO THE PRESIDENT
Michael Kratsios
Saved 2026-07-22
SP2BMHRK
webpage
Saved 2026-07-22
NFY63EP3
webpage
PhD
Saved 2026-07-22
687KYGFP
journalArticle
Lukas Röseler, Leonard Kaiser, Christopher Doetsch, Noah Klett, Christian Seida, Astrid Schütz, Balazs Aczel, Nadia Adelina et al.
2024 · Journal of Open Psychology Data
Saved 2026-07-22
A9N8Z9UR
webpage
Saved 2026-07-22
Q4NETH4D
webpage
Saved 2026-07-22
52T3V9DG
preprint
Joseba Fernandez de Landa, Carla Perez-Almendros, Jose Camacho-Collados
2026
Saved 2026-07-22
LLMs have been showing limitations when it comes to cultural coverage and competence, and in some cases show regional biases such as amplifying Western and Anglocentric viewpoints. While there have been works analysing the cultural capabilities of LLMs, there has not been specific work on highlighting LLM regional preferences when it comes to cultural-related questions. In this work, we propose a new dataset based on a comprehensive taxonomy of Culture-Related Open Questions (CROQ). The results show that, contrary to previous cultural bias work, LLMs show a clear tendency towards countries such as Japan. Moveover, our results show that when prompting in languages such as English or other high-resource ones, LLMs tend to provide more diverse outputs and show less inclinations towards answering questions highlighting countries for which the input language is an official language. Finally, we also investigate at which point of LLM training this cultural bias emerges, with our results suggesting that the first clear signs appear after supervised fine-tuning, and not during pre-training.
2GY4ANP8
webpage
Micah Carroll
Saved 2026-07-21
RNH8J3D3
blogPost
Alex Koren
2016
Saved 2026-07-21
I get asked a lot how to apply for the Thiel Fellowship and it usually boils down to two questions:
8X4EFBMR
webpage
Saved 2026-07-21
B46WEUWZ
preprint
Brett Reynolds
2026
Saved 2026-07-20
Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an instruction, refused appropriately, complied with a policy, resisted an embedded command, or misreported progress in an agentic task. Existing benchmarks often compress these distinctions into pass/fail labels, obscuring whether failures arise from capability limits, policy ambiguity, instruction conflict, scaffold failure, or unstable evaluator judgments. This paper introduces adversarial pragmatics as a benchmark and annotation protocol for evaluating model behaviour under instruction conflict, embedded commands, quotation, scope ambiguity, deixis, indirect speech acts, and multi-turn agent transcripts. The contribution is empirical and methodological: a linguistically controlled taxonomy, an 18-item seed benchmark with validator-enforced metadata, a 54-row local seed pilot, an expert-evaluation protocol distinguishing task success, policy compliance, safety risk, refusal outcome, and evaluator confidence, and metrics for judge validity, diagnostic ambiguity, and taxonomy drift. The benchmark treats labels as inference licenses: it tests whether safety-relevant categories project across paraphrase, wrapper, model, and judge condition. In the pilot, a rubric-aided LLM judge graded its own outputs with expected-behaviour fields visible and still missed the safety-relevant minority classes.
RSHG3PFQ
preprint
Brett Reynolds
2026
Saved 2026-07-20
Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an instruction, refused appropriately, complied with a policy, resisted an embedded command, or misreported progress in an agentic task. Existing benchmarks often compress these distinctions into pass/fail labels, obscuring whether failures arise from capability limits, policy ambiguity, instruction conflict, scaffold failure, or unstable evaluator judgments. This paper introduces adversarial pragmatics as a benchmark and annotation protocol for evaluating model behaviour under instruction conflict, embedded commands, quotation, scope ambiguity, deixis, indirect speech acts, and multi-turn agent transcripts. The contribution is empirical and methodological: a linguistically controlled taxonomy, an 18-item seed benchmark with validator-enforced metadata, a 54-row local seed pilot, an expert-evaluation protocol distinguishing task success, policy compliance, safety risk, refusal outcome, and evaluator confidence, and metrics for judge validity, diagnostic ambiguity, and taxonomy drift. The benchmark treats labels as inference licenses: it tests whether safety-relevant categories project across paraphrase, wrapper, model, and judge condition. In the pilot, a rubric-aided LLM judge graded its own outputs with expected-behaviour fields visible and still missed the safety-relevant minority classes.
BZ74CTB8
preprint
Otto Jespersen, Brett Reynolds, Peter Evans
2025
Saved 2026-07-20
This volume presents a new edition of Otto Jespersen's landmark 1917 study of negation in English and other languages, primarily Germanic and Romance. While best known for describing what would later be called “Jespersen's Cycle'”, this work offers far more: a comprehensive analysis of negative expressions, their forms, functions, and historical development. The book examines topics ranging from negative prefixes to the distinction between special and nexal negation, supported by Jespersen's characteristically rich collection of authentic examples. This edition features an extensive new introduction by Olli O. Silvennoinen that situates Jespersen's work in its historical and intellectual context while highlighting its continued relevance to contemporary linguistics. The main text has been entirely re-typeset to enhance readability, with examples presented in modern numbered format and Leipzig-style glosses added for non-English examples. Where possible, hyperlinks to source materials have been provided, making this classic work more accessible than ever for modern scholars and students of linguistics.
6R7X4NBG
forumPost
Zohar Atkins [@ZoharAtkins]
2026
Saved 2026-07-20
E5HTGGDS
webpage
Niall Ferguson
2026
Saved 2026-07-20
The tools that once exposed and debunked Holocaust denial are powerless against AI and the algorithm. Niall Ferguson and John-Clark Levin ask: Is there a remedy?
S7IY94V2
webpage
Saved 2026-07-20
8ZVFGMHC
webpage
Saved 2026-07-20
G7G6N7CY
webpage
2026
Saved 2026-07-20
Codex (wife) took custody of the kids (dreams and whimsy) and now i am in a social club at 1:30 am confronting my thoughts under the influence of tequila.