Andrew Wilkinson Launches 'Deep Personality' After Vibe Coding Controversy
Is this a scandal?
No longer — the story has resolved. Noise 2/100, cooling down, across 1 source.
Regulatory scrutiny regarding 'AI medical devices' or diagnostic software is likely to increase as similar DIY health apps proliferate. Wilkinson will likely face pushback from medical boards or consumer protection groups if the app is perceived as providing unlicensed clinical advice.
Noise 2/100 — louder than 95% of tracked AI controversies.
Why it matters
Demonstrates how rapid AI prototyping enables non-experts to deploy high-stakes diagnostic tools, bypassing traditional clinical validation and regulatory oversight.
Key points
- Deep Personality screens for 30+ mental health conditions using 28 research-backed assessments in under 40 minutes.
- Andrew Wilkinson built the entire application using Anthropic's Claude Code via natural language prompting.
- The app provides diagnostic insights and treatment roadmaps without licensed clinician oversight or regulatory approval.
- Development exemplifies 'vibe coding' trend enabling rapid deployment of complex software by non-programmers.
- Launch raises safety concerns about unvalidated AI tools performing high-stakes mental health evaluations.
- No evidence indicates Anthropic reviewed or approved the application's medical claims or methodology.
The story
Entrepreneur Andrew Wilkinson has launched Deep Personality, an AI-powered web application that screens users for over 30 mental health conditions using Claude Code. The tool consolidates 28 clinically validated assessments into a 40-minute evaluation, providing personalized treatment roadmaps without professional supervision. Wilkinson developed the platform through "vibe coding," a method where natural language prompts generate functional software with minimal human programming. While marketed as science-backed, the application operates outside standard healthcare regulations despite offering diagnostic-level insights. Industry observers note this represents a growing trend of AI-generated health tools entering the market rapidly. Critics argue such applications risk misdiagnosis and lack accountability mechanisms inherent in traditional telehealth. Anthropic has not commented on third-party medical use of its models. The launch highlights tensions between AI democratization and patient safety in unregulated digital health markets.
Who's involved
Likely to express concerns over the lack of clinical validation and the risks of self-diagnosis via unvetted AI tools.
Argues that AI-driven screening democratizes access to mental health insights for those who cannot afford traditional therapy.
Provider of the underlying coding technology used to build the application's interface and logic.
Most contested claim
Deep Personality provides valid mental health screening and clinical-grade insights through AI aggregation.
Biggest open question
Whether the specific combination and AI-interpretation of these tests have been independently validated as a unified diagnostic instrument.
Read the full story
How we got here
The deployment of AI-generated software in regulated domains follows a recurring pattern where technical capability outpaces domain-specific validation protocols. Historically, the digitization of clinical psychometrics required rigorous translation studies to ensure that changing the medium of administration (e.g., from paper to digital) did not alter psychometric properties. Current precedents show that when non-clinical actors leverage generative AI to assemble diagnostic workflows, the resulting systems often inherit the validity of individual components without establishing the validity of the composite interpretation layer. This mirrors earlier controversies in direct-to-consumer genetic testing and wellness apps, where the aggregation of scientific data points created an impression of clinical authority absent regulatory clearance. The pattern suggests that 'vibe coding' or rapid AI prototyping lowers the barrier to entry for building functional artifacts but does not reduce the epistemic requirements for claiming diagnostic utility. Consequently, industry friction consistently arises at the intersection of open-source model capabilities and closed-loop professional liability standards.
The full story
On February 4, 2026, entrepreneur Andrew Wilkinson publicly launched 'Deep Personality,' a web application designed to screen users for over 30 mental health conditions and provide personalized roadmaps for seeking help. The launch followed a period of rapid prototyping in January 2026, during which Wilkinson utilized AI coding assistants to build the platform. According to Wilkinson, the application was constructed using Claude Code to implement the interface and logic, while ChatGPT was employed to identify and aggregate professional psychological tests. Wilkinson described the tool as performing a 'deep analysis' on personality and relationships to identify potential problems, positioning it as a science-backed alternative to traditional assessments.
The product's marketing materials state that Deep Personality consolidates approximately 28 to 30 research-backed assessments covering personality traits, attachment styles, neurodiversity, and mental health screening into a single session lasting under 40 minutes. Wilkinson emphasized the speed of development and low cost enabled by AI tools, framing the project as a demonstration of how non-experts can now deploy complex diagnostic interfaces. In public statements, he argued that such tools democratize access to mental health insights for individuals unable to afford traditional therapy or clinical evaluations.
However, the launch immediately drew scrutiny regarding clinical validation and safety. While Wilkinson asserts the app uses 'clinically validated' tests, critics from the mental health profession have raised concerns about the absence of independent clinical trials for the aggregated system itself. The core controversy centers on whether combining validated individual instruments into an AI-interpreted composite constitutes a new, unvalidated diagnostic tool. Mental health professionals argue that self-diagnosis via unvetted AI platforms carries risks of misinterpretation, false positives, and inadequate crisis management, distinguishing between administering a test and providing a clinical synthesis.
Anthropic, the provider of Claude Code, remains a neutral party in this specific dispute; their technology facilitated the software's creation but does not imply endorsement of its medical efficacy. The timeline indicates a compressed development cycle where the transition from identifying tests to launching a public-facing mental health screening tool occurred within weeks. This sequence highlights the tension between technical feasibility—building a functional testing interface—and clinical readiness, which typically requires longitudinal validation. As of the current resolution state, no formal regulatory action has been cited in the available sources, but the discourse has established a clear dichotomy between technological democratization and professional standards of care.
What's confirmed, what's disputed
- ConfirmedDeep Personality screens users across 30+ mental health conditions and provides a help roadmap.
- ConfirmedThe app consolidates 28 research-backed assessments into a session lasting under 40 minutes.
- ConfirmedWilkinson used Claude Code to build the app and ChatGPT to identify professional psychological tests.
- ConfirmedThe app functions as an AI relationship counselor and personality analyzer.
- DisputedDeep Personality consolidates clinically validated personality tests.
The strongest case each way
Aggregating validated tests does not validate the aggregate; without clinical norms for the combined output, the tool risks generating misleading diagnostic signals that lay users may treat as medical advice.
Democratizing access to established psychological frameworks via low-cost AI tools bridges the gap for underserved populations who otherwise have zero access to mental health screening.
Times this happened before
- BetterHelp Data Privacy & Licensing Controversy · 2024FTC settlement and revised licensing disclosures
- National Eating Disorders Association (NEDA) Tessa Chatbot Shutdown · 2024Immediate suspension after harmful advice reports
What's at stake
The primary stakeholders are lay users seeking mental health guidance, who face potential harm from false positives/negatives in an unvalidated composite assessment. Mental health professionals risk erosion of trust if AI tools conflate test administration with clinical diagnosis. Conversely, proponents argue the stakes include continued exclusion of low-income individuals from any form of structured psychological insight. The magnitude is currently diffuse, affecting early adopters of AI health tools, but sets a precedent for mass-market AI diagnostics. If such tools gain traction without validation, the cumulative risk involves systemic misallocation of mental health resources and user distress. The defender's stake is the preservation of innovation space for non-traditional health tech entrants.
What we still don't know
- Whether the specific combination and AI-interpretation of these tests have been independently validated as a unified diagnostic instrument.
Noise Level
The timeline
- Last month
Initial Experimentation
Wilkinson uses ChatGPT to identify professional psychological tests and Claude Code to build a unified testing interface.
Deep Personality Launch
Wilkinson publicly announces the app on social media, emphasizing its speed of development and low cost.
The full record
Sources & methodology
- Andrew Wilkinson — x.com · located later (2026-07-30)
- Transcript: 'Opus 4.5 Changed How Andrew Wilkinson Works ... — every.to · located later (2026-07-30)
- A custom email client he built by handing Claude Code his ... — x.com · located later (2026-07-30)
- Opus 4.5 Changed How Andrew Wilkinson Works and Lives — every.to · located later (2026-07-30)
- Deep Personality: Science-backed personality insights for ... — producthunt.com · located later (2026-07-30)
The records from this story's original coverage were pruned, so items marked located later were found by searching for it afterwards. The summary above has since been rewritten to take them into account — it is not the text first published. How we score →
Where the sources disagree
In dispute Deep Personality provides valid mental health screening and clinical-grade insights through AI aggregation.
Established Deep Personality administers multiple existing research-backed assessments via an AI-coded interface, but lacks evidence of validation for its composite diagnostic output.
What's being under-reported
Missing perspective: End-user experience data and clinical psychologist reviews of the actual output quality. Current coverage focuses on development narrative and marketing claims, lacking empirical assessment of whether the AI's interpretations are clinically coherent or dangerous. This gap matters because the true risk profile depends on output quality, not just input validity.
Who changed their mind, and why
- Andrew WilkinsonMaintained consistent positioning of the tool as a democratizing force and technical showcase, emphasizing speed and accessibility over clinical endorsement. (was: Initial experimentation phase focused on technical feasibility of AI-coded psychometrics.)
- Mental Health ProfessionalsShifted from general AI skepticism to specific critique of unvalidated composite diagnostics following the public launch claims.
The forecast
Regulatory scrutiny regarding 'AI medical devices' or diagnostic software is likely to increase as similar DIY health apps proliferate. Wilkinson will likely face pushback from medical boards or consumer protection groups if the app is perceived as providing unlicensed clinical advice.
Forecast, not fact — an editorial estimate we score when this resolves.
That's the complete picture as of — nothing more to know right now. We'll update this page the moment it changes.
Join the Discussion
Discuss this story
Community comments coming in a future update
Be the first to share your perspective. Subscribe to comment.