Named after the hundred-eyed watchman of Greek myth, Argus watches the education landscape: spotting new opportunities, pressure-testing the ventures we're building, and tracing every read back to the real-world signals behind it.
The evidence library: the raw signals the pipeline is watching across the education ecosystem. Every idea is built from these.
Improving diversity in U.S. Alzheimers disease (AD) research is a pressing need. By 2050, Hispanic and Latino Americans will comprise 30% of the population. Hispanics are 1.5 times more likely and Blacks are twice as likely to develop AD compared to Whites, yet both remain vastly underrepresented in clinical trials research. Aging and AD research mentorship of underrepresented STEM undergraduates is designed to promote entry into related professions by students committed to decreasing disparities in AD research participation and clinical care. The NIA-funded MADURA program recruited 93 students from backgrounds historically underrepresented in STEM majors and/or from NIH-defined disadvantaged backgrounds. Trainees were placed in aging/AD research labs and received weekly training and mentorship from faculty research PIs and other types of supervisors (postdoctoral researchers, graduate students, research assistant staff...) Our study examined student ratings of the program and mentor b
Objective Structured Clinical Examinations (OSCEs) are widely used to assess medical students clinical skills, including non-technical abilities such as communication and empathy. However, the potential influence of individual psychological traits--such as personality dimensions, empathy, and stress-related mindset--on OSCE performance remains understudied. This study investigated associations between personality traits, empathy levels, stress mindsets, and performance in OSCEs among medical students. An online questionnaire (including the Big Five Personality Traits Inventory 2, the Jefferson Scale of Physician Empathy (Medical Student version), the Growth Mindset Scale, the Stress Mindset Measure) was provided to all fifth-year medical students enrolled at the Universite Paris Cite for six weeks before undertaking graduation summative OSCEs. Their scores were correlated with OSCE performance using Spearmans correlation and linear regression analyses. A total of 99 questionnaires were
BackgroundPeriodontal disease (PD) and diabetes mellitus (DM) have a well-established bidirectional relationship, affecting glycaemic control and chronic disease outcomes. However, the extent to which medical training supports physician awareness of this association remains unclear especially in resource-limited settings. ObjectiveTo assess exposure to oral health education and to identify predictors of awareness of PD-DM association among physicians. MethodsA cross-sectional study was conducted among 146 physicians managing diabetic patients at a tertiary teaching hospital in Ghana. A structured questionnaire assessed exposure to oral health education, periodontal disease knowledge (score range 0-5), and awareness using a 5-item Likert scale (score range 5-25). Multivariable linear regression identified predictors of awareness. ResultsAlthough 62.1% reported exposure to oral health content during undergraduate training, 59.2% rated its quality as poor. Mean awareness score was 20.6 (S
BackgroundHigh workload among healthcare workers has increasingly been correlated with poor patient outcomes, inefficient operational and financial outcomes, and burnout. Despite growing literature exploring causes of attending physician workload, there is limited understanding of trainee-specific measures. ObjectiveWe aimed to characterize elements contributing to trainee workload and perceived challenges and satisfiers to the trainee workday as a foundation for better understanding and measuring trainee work experience. MethodsInternal Medicine and Medicine-Pediatrics residents at an academic medical center were invited to participate in focus groups discussing contributors to inpatient workload and work experience between March and April 2024. A qualitative content analysis identified key metrics of trainee workload and work experience, which were then consolidated into overarching domains. A structured, multi-round rating process ranked the perceived relevance of each metric. Resul
IntroductionEvaluate large language models (LLMs) for scoring medical student essays, and compare various prompting techniques and models. MethodsOpenAI GPT scored 51 medical student reflection essays (15 real, 36 fabricated) using a previously-reported 6-point rubric (April-May 2025). We compared 29 prompt-model conditions by systematically varying the LLM prompts (including the persona, scoring rubric, few-shot learning [exemplars], chain-of-thought reasoning, and temperature), fine-tuning, and model (including GPT-4.1, GPT-4.1-mini, GPT-o4-mini, and GPT-4-Turbo). Outcomes were accuracy (compared with human raters, measured using single-score intraclass correlation coefficient [ICC] and mean absolute difference [MAD; zero indicates perfect agreement]), within-condition reproducibility, and cost. ResultsAcross all conditions, it took mean (SD) 3.73 (3.12) seconds to score 1 essay. The cost to score 100 essays was USD $0.04 for GPT-4.1-mini, $0.21 for GPT-4.1, $0.57 for GPT-4.1 with 3
In recent years, progress in medical informatics and machine learning has been accelerated by the availability of openly accessible benchmark datasets. However, patient-level electronic medical record (EMR) data are rarely available for teaching or methodological development due to privacy, governance, and re-identification risks. This has limited reproducibility, transparency, and hands-on training in cardiovascular risk modelling. Here we introduce PRIME-CVD, a parametrically rendered informatics medical environment designed explicitly for medical education. PRIME-CVD comprises two openly accessible synthetic data assets representing a cohort of 50,000 adults undergoing primary prevention for cardiovascular disease. The datasets are generated entirely from a user-specified causal directed acyclic graph parameterised using publicly available Australian population statistics and published epidemiologic effect estimates, rather than from patient-level EMR data or trained generative mode