Named after the hundred-eyed watchman of Greek myth, Argus watches the education landscape: spotting new opportunities, pressure-testing the ventures we're building, and tracing every read back to the real-world signals behind it.
The evidence library: the raw signals the pipeline is watching across the education ecosystem. Every idea is built from these.
arXiv:2606.22689v2 Announce Type: replace Abstract: As AI-generated content (e.g., "slop") becomes more prevalent online, people are developing strategies to attempt to identify it (or, conversely, to gain confidence that something is not AI-generated). What strategies are people using, and how are they changing over time as generative AI models themselves change? In this work, we catalog and analyze 2 years and 8 months of the AI detection strategies discussed by users of two popular Reddit communities (r/isthisAI and r/RealOrAI) that use the wisdom of crowds to identify AI-generated media. Through a mixed-method analysis of 13,098 posts and 222,060 comments within these communities, we catalog and analyze the prevalence of 12 AI-detection strategies, including examining fine-grained physical details, recognizing trends in AI-created content, and the assumptions people make about what models are capable of producing. Furthermore, we find that these strategies and mental models shift o
arXiv:2604.24155v3 Announce Type: replace Abstract: The project of aligning machine behavior with human values raises a basic problem: whose moral expectations should guide AI decision-making? Much alignment research assumes that the appropriate benchmark is how humans themselves would act in a given situation. Studies of agent-type value forks challenge this assumption by showing that people do not always judge humans and AI systems identically.This paper extends that challenge by examining two further possibilities: first, that evaluations of AI behavior change when its human origins are made visible; and second, that people judge the humans who program AI systems differently from either the machines or the human actors they are compared against. An experiment with 1,002 U.S. adults measured moral judgments in a runaway mine train scenario, varying the subject of evaluation across four conditions: a repairman, a repair robot, a repair robot programmed by company engineers, and compan
arXiv:2601.14264v2 Announce Type: replace Abstract: Large language models (LLMs) act as digital twins for human respondents, yet their psychometric comparability remains uncertain. We propose a construct validity framework spanning construct representation and the nomothetic span, benchmarking models against human gold standards. Across studies, digital twins achieved high aggregate-level accuracy and profile correlations, but showed attenuated item-level correlations. In word association tests, LLM networks exhibited humanlike small-world structure and theory-consistent communities, yet diverged lexically and in local structure. In decision-making and contextualized tasks, they under-reproduced heuristic biases, demonstrating normative rationality, compressed variance, and limited temporal sensitivity. Feature-rich and trait relevant conditioning improved Big Five personality prediction and nomothetic-span alignment, but network invariance remained limited, with partial configural sol
arXiv:2606.28277v1 Announce Type: cross Abstract: Artificial intelligence is driving a revolution in scientific discovery, accelerating everything from hypothesis generation to mathematical theorem proving. However, this rapid acceleration is creating a systemic challenge: traditional human peer review cannot scale to match the influx of AI-assisted science. Ultimately, to resolve this tension, we must also deploy AI to accelerate the verification and review process itself. To frame the discussion around this transition, we propose a taxonomy consisting of four progressive levels of AI-human collaboration in scientific evaluation, and discuss various trade-offs involved with each. As a step toward this future, we introduce the Paper Assistant Tool (PAT), an agentic AI framework built for deep scientific review and verification. PAT ingests full scientific manuscripts and produces a comprehensive evaluation, checking theoretical results, validating experiments, suggesting improvements,
arXiv:2606.28186v1 Announce Type: cross Abstract: Predicting human item difficulty is central to educational assessment, where reliable estimates support fairness and effective test construction. Existing methods often depend on costly human calibration or item-level textual representations, providing limited evidence about the cognitive processes that make items difficult. We argue that difficulty should be viewed not only as a property of item text, but also as an observable consequence of the problem-solving burden an item induces. Large Reasoning Models (LRMs) offer scalable process evidence through reasoning traces, but such evidence must be structured to support interpretable modeling. To this end, we introduce Epi2Diff (Episode to Difficulty), a framework that maps LRM reasoning traces into cognitively grounded episode sequences. These episodes group trace segments into functional problem-solving states, enabling difficulty to be modeled through reasoning scale, effort allocatio
arXiv:2606.27951v1 Announce Type: new Abstract: AI agents are promising tools that can act as flexible behavioral nudges to enhance human cooperation in addressing large-scale societal problems. However, evidence on whether AI agents can effectively boost cooperation remains mixed. We recruited 1,283 participants to play iterated Collective Risk Games in small groups, testing whether AI assistants could nudge participants toward cooperation. By using persuasive framing personalized to each player's Social Value Orientation profile, the AI interventions significantly increased contributions and group success rates. These cooperative effects were short-lived, however, fading after the first few rounds. Strikingly, when the AI treatments were reconfigured to promote selfish behavior through exculpatory framing, the negative effects on contributions and group success were larger and substantially more persistent, particularly for personalized interventions. This asymmetry between prosocial
arXiv:2606.27689v1 Announce Type: new Abstract: AI-driven deception mechanisms are increasingly prevalent in digital games, yet the direction and magnitude of their effects on player experience remain contested. Existing research has not sufficiently disentangled designer-intended deception intensity from players' actual perception of deception, and most prior work relies on low-ecological-validity experiments or cross-sectional surveys. The present study aims to independently examine the causal effects of design deception intensity (DDI) and player deception awareness (PDA) on player ratings within a naturalistic gaming environment, and to investigate the moderating role of player experience. Leveraging the 54 version updates of Baldur's Gate 3 between 2019 and 2025 as a quasi-natural experiment, it collected all English-language Steam reviews posted within 1 to 28 days following each update, and constructed a player-version two-way fixed effects panel dataset. DDI was coded by human
There is a squeaky old merry-go-round in my neighborhood that my own children play on from time to time. Years of kids riding on it have loosened its joints so it spins more freely and quickly.
You’ll often hear two words come up in advising sessions as students look ahead to college: match and fit. They sound interchangeable, but they’re not.
Article URL: https://digilent.com/blog/why-communication-might-be-the-missing-link-in-engineering-education/ Comments URL: https://news.ycombinator.com/item?id=49076962 Points: 3 # Comments: 0
Lakshmi Halasyamani, chief clinical officer at Endeavor Health, thinks value-based care isn’t really about payment models. To her, this work centers on understanding patients’ clinical, social and financial circumstances well enough to make care plans that actually work for their lives. The post Why Endeavor Health Refuses to Silo Its Value-Based Care Strategy appeared first on MedCity News .
Argenx’s first acquisition is the buyout of Forte Biosciences, a company with a lead drug that has early clinical validation in celiac disease and vitiligo. Argenx executives say this antibody complements its blockbuster product Vyvgart, offering a different mechanism of action but similar pipeline-in-a-product potential. The post Argenx’s $2B Forte Bio Acquisition Brings Autoimmune Playbook to Prevalent Disorders appeared first on MedCity News .
New Options Emerge for Disenrolled Howard Students Joshua.Bay Mon, 07/27/2026 - 04:58 PM SUNY, CUNY and U of the District of Columbia are offering support and enrollment opportunities to those affected by Howard’s decision to unenroll 502 incoming students. Byline(s) Joshua Bay
Article URL: https://builders.ramp.com/post/thompson-sampling-model-routing Comments URL: https://news.ycombinator.com/item?id=49074574 Points: 1 # Comments: 0
During Pride month last year, almost 70,000 LGBTQ+ young people sought specialized help from 988, the national suicide and crisis hotline. That was the last time they were able to do so. As of today, it’s been a year since the 988 hotline’s LGBTQ+ youth services shut down. The Trump administration says it is working […]
Article URL: https://blainehansen.me/post/learning-is-for-students-not-llms/ Comments URL: https://news.ycombinator.com/item?id=49073349 Points: 7 # Comments: 1
Yale, Columbia, Brown and a raft of other institutions warned of damages to the national research system from “draconian” research funding cuts.
A new legal challenge to Texas’ school funding system renews questions about what methods state leaders should use to ensure students in low-income districts and those attending campuses in more affluent communities receive an education on a level playing field. The Midland school district voted to sue Education Commissioner Mike Morath and state leaders this […]
Universities produce an enormous amount of visual content. Faculty members create lecture presentations and course materials. Marketing teams produce event promotions and recruitment campaigns. Students generate slides, videos and graphics for coursework. But keeping all of those efforts coordinated while maintaining quality and brand consistency can be difficult. Canva for Campus is designed to simplify that process by providing a shared platform where students, faculty and staff can design, collaborate and publish visual materials without needing specialized design software. Canva for…
In 2020, voters in California’s Alameda County made a remarkable decision. Rather than accept generations of underinvestment in young children, they chose to build something families had never truly had: an early childhood system designed to work. Parents, early educators, childcare providers, labor leaders, business leaders, advocates and community organizations came together around a shared […]
Students helped design the A.I.-powered creation as a young female with dark hair and an upbeat personality. Then came the outrage. The post Sally the robot was coming to a New York school. Then the plug was pulled. appeared first on District Administration .
When staff are overextended and patients face delays or abandon care, the issue is no longer a narrow administrative burden. It becomes a business, financial, and access challenge that demands a different operating model. The post Prior Authorization Is Draining Revenue: Why Automation Has Become a Strategic Imperative appeared first on MedCity News .
The question is not whether the technology is impressive. The question is whether it shows up at the right time, with the right information, inside the workflow where the decision is being made. Anything else is just another report. The post How AI Inside Clinical Workflows Is Unlocking Patient Throughput appeared first on MedCity News .
Texas may expand use of the Classical Learning Test despite a new report finding less evidence its scores predict college success than the SAT or ACT. The post Universities can use conservative-favored entrance exam as Texas eyes wider use, agency chief says appeared first on District Administration .
OKLAHOMA CITY — State leaders say they’re still pursuing solutions to reverse a yearslong decline in traditionally trained teachers entering the classroom, all while emergency certified educators continue to fill essential teaching roles. The number of graduates completing Oklahoma teacher preparation programs, including college degrees in education, has fallen by about 40% since 2013, state […]
[Sponsored] On a recent webinar sponsored by Verato, panelists from SCAN Health Group and the Alliance of Community Health Plans discussed how their organizations are meeting the moment. The post AI Readiness Starts with Solving Healthcare’s Data Fragmentation Problem appeared first on MedCity News .
As patients increasingly bring AI chatbot advice into exam rooms, where does liability land when that advice goes wrong? Health law attorney Meghan O’Connor said the answer still isn’t clear, but silence is the riskiest response for providers. The post What Happens When Patients Trust AI Over Their Doctor? appeared first on MedCity News .
A few years ago, education experts noticed that scores were climbing upwards on a number of Advanced Placement tests. The increases, measured on some of the program’s most popular assessments, were as significant as they were unexpected. In AP Chemistry, the share of students scoring a 3, 4 or 5 (generally interpreted as a passing […]
Anti-bullying and social-emotional learning haven’t reduced school shootings. Minimizing public attention on shooters might.
Professional development is one of a school district's most powerful strategies for improving student outcomes. Yet, educators often leave workshops feeling inspired in the moment, but take little action in the classroom. The problem isn’t that teachers don’t want to learn, but that professional development is treated as an event, rather than an ongoing process.
Scientists rarely announce that one of their long-held ideas has failed. But that is exactly what Ron Avi Astor, one of the nation’s leading scholars of school violence and a professor at the University of California, Los Angeles, said before fellow researchers at the annual meeting in April of the American Educational Research Association (AERA). […] The post A leading school shootings researcher says he was wrong appeared first on The Hechinger Report .
What happens when you build EdTech features by actually listening to students? Kyron Learning partnered with the AIMS Collaboratory and the Gates Foundation to study motivation and engagement among English language learners in middle school math, and the findings drove real product changes. This piece traces how closed captioning, highlighted transcripts, and clickable definitions emerged from classroom observation rather than assumption, offering a model for what participant-driven, learner-centered product development can look like at its best. The post How Student Feedback Shaped New Features for English Language Learners in Math appeared first on Getting Smart .
LIFT, or Literacy Intervention for Teens, is built to help older struggling readers succeed.
Among college students, 37 percent report moderate to severe depressive symptoms and 33 percent report moderate to severe anxiety. More than three-quarters (76 percent) of college students report moderate or high stress levels, and almost one-quarter (20.6 percent) experience significant psychological distress while on campus. The post Strengthening campus pathways to student mental health care appeared first on eCampus News .
Iowa State Ordered to Pay Former Professor $2.8M in Equal Pay Lawsuit Emma Whitford Mon, 07/27/2026 - 03:00 AM Byline(s) Emma Whitford
Workforce Pell Is Live. How’s It Going for States? Sara Weissman Mon, 07/27/2026 - 03:00 AM As states work to implement the new policy, only a trickle of programs have qualified. But experts say it’s too early to tell Workforce Pell’s full impact. Byline(s) Sara Weissman
Reed College Settles Federal Antisemitism Complaint kathryn.palmer… Mon, 07/27/2026 - 03:00 AM Byline(s) Kathryn Palmer
NCTE Dissolves 77-Year-Old Composition Association Emma Whitford Mon, 07/27/2026 - 03:00 AM The National Council of Teachers of English’s decision was met with swift backlash from Conference on College Composition and Communication members, who say their organization is vital to advancing composition research and improving teaching. Byline(s) Emma Whitford
Can Outcomes-Based Financing Help Solve the Loan Crisis? jessica.blake@… Mon, 07/27/2026 - 03:00 AM In an interview with Inside Higher Ed , Ethan Pollack, a senior director at Jobs for the Future, discusses how these loans work and why they might help address today’s challenges. Byline(s) Jessica Blake
Building a Road Map for First-Gen Students Joshua.Bay Mon, 07/27/2026 - 03:00 AM The University of Houston–Downtown’s First-Generation Gator Program combines mentoring, scholarships and family engagement to help students navigate college. Byline(s) Joshua Bay
Dueling Mandates Sara Brady Mon, 07/27/2026 - 03:00 AM Serve the community vs. run the college like a business. Byline(s) Matt Reed
Howard Reinstates Classes for Dozens After Unenrolling Hundreds sara.custer@in… Mon, 07/27/2026 - 03:00 AM Byline(s) Sara Custer
Extortion and Appeasement at Northwestern Sara Brady Mon, 07/27/2026 - 03:00 AM There is so much government repression happening that I cannot begin to summarize it. Byline(s) John K. Wilson
We’re rounding up recent stories, from civil rights groups slamming a regulatory change to a ruling against the administration’s approach to grant cuts.
In court documents, the National Institutes of Health, the National Science Foundation and others shared which words they used to target grants.
We’re rounding up last week’s news, from absenteeism recovery data to a slew of new legislative proposals.
The superintendents group said allowing public benefits usage to be considered in noncitizens’ applications could have a “chilling effect” on schools.
In too many classrooms in America, the scene is the same: children staring at tablets, headphones on, silent. Some are working through reading programs. Some have already finished and found their way to something else entirely. The teacher circulates. The education technology industry calls this personalized learning. Somewhere, a dashboard is registering all this as […] The post OPINION: Walking away from education technology isn’t the answer to the latest backlash appeared first on The Hechinger Report .
arXiv:2607.18292v2 Announce Type: replace-cross Abstract: As language models scale, answers start truer but degrade faster: scaling buys capability but erodes reliability. The knowledge-gap account -- more data, retrieval, or scale -- misses an auto-regressive risk residual that increases with scale: the model commits to a low-probability token, conditions on it as established, and snowballs. We track this through per-position disagreement $\delta = \log p_M - \log p_O$ against a stronger same-family oracle, whose second moment splits exactly into bias$^2$ $\mathrm{KL}(p_M \,\|\, p_O)^2$ and risk $\mathrm{Var}[\delta]$. Across three model families, we present four findings: (i) under scaling, the knowledge gap falls up to $7\times$ while knowledge degradation grows up to $39\times$; (ii) at a fabrication, felt uncertainty $H(p_M)$ relaxes quickly while oracle-referenced risk persists up to $23\times$ longer, leaving a confident-but-precarious risk regime that bridges consecutive fabric
arXiv:2606.20295v2 Announce Type: replace-cross Abstract: Large model inference optimization serves as a key foundation for supporting the scalable, low-cost, and highly stable operation of large model services. Centered on token-oriented inference optimization technology, this paper proposes for the first time a four-layer technical architecture consisting of Multi-model Fusion, Model Optimization, Compute-Model Fusion, and Compute-Network-Model Fusion. It systematically reviews the key technologies and current industry status across these four levels and analyzes the application value of related technologies in real-world business scenarios. This paper provides a practical technical path for reducing token production costs, improving token service efficiency, ensuring the stability of token supply, and driving the transition of large model services from being merely callable to being operable.