Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review.
How this is built →
2Distinct papers
42Unique collaborators
2/2Semantic Scholar citation coverage
Publication span: 2026. Corpus fetch span: 2026.
Identity provenance
Provider IDs
- Semantic Scholar:
2263519562
ORCID evidence
No valid ORCID is stored.
Observed aliases (1)
- Elaine Lau (semantic scholar, provider refresh)
Topics and outcomes in this view
Assessment themes
- Human Ai Collab: 2 papers
- Productivity: 2 papers
Claim outcomes
- Other: 1 paper
- Research Productivity: 1 paper
- Decision Quality: 1 paper
- Output Quality: 1 paper
- Error Rate: 1 paper
- Task Completion Time: 1 paper
Papers in the Semantic Scholar view
Latest stored Semantic Scholar author observations only. Citation counts below are from the same provider and are not combined with other services.
Scroll the table horizontally to see every column.
| Paper | Author evidence | Date | Provider citations |
|---|---|---|---|
| A high-fidelity benchmark of junior investment-banker workflows shows frontier AI still far from ready for delegation: the best model fails roughly half of task criteria and produces no client-ready outputs in banker assessments; failures cluster on cross-document consistency and tool-based data retrieval.arxiv | Elaine Lau provider id |
2026-04-13 | 1 |
| Large language models struggle to predict experimental outcomes reliably: they reach only 14–26% accuracy—comparable to human experts—but cannot tell when their predictions are trustworthy; human experts, by contrast, are well calibrated and markedly better at identifying which outcomes can be predicted without physical tests.arxiv | Elaine Lau provider id |
2026-04-12 | 1 |
Citation observation summary
Semantic Scholar supplied counts for 2 of 2 papers in this view; 0 are missing. The observed paper counts sum to 2 cumulative citations. This is a coverage summary, not an author score or h-index.