Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review.
How this is built →
2Distinct papers
349Unique collaborators
2/2Semantic Scholar citation coverage
Publication span: 2026. Corpus fetch span: 2026.
Identity provenance
Provider IDs
- Semantic Scholar:
2325951260
ORCID evidence
No valid ORCID is stored.
Observed aliases (1)
- Hanwen Xing (semantic scholar, provider refresh)
Topics and outcomes in this view
Assessment themes
- Adoption: 2 papers
- Human Ai Collab: 2 papers
- Productivity: 2 papers
- Innovation: 1 paper
- Org Design: 1 paper
- Skills Training: 1 paper
Claim outcomes
- Other: 2 papers
- Adoption Rate: 1 paper
- Firm Revenue: 1 paper
- Fiscal And Macroeconomic: 1 paper
- Market Structure: 1 paper
- Output Quality: 1 paper
- Regulatory Compliance: 1 paper
- Task Completion Time: 1 paper
Papers in the Semantic Scholar view
Latest stored Semantic Scholar author observations only. Citation counts below are from the same provider and are not combined with other services.
Scroll the table horizontally to see every column.
| Paper | Author evidence | Date | Provider citations |
|---|---|---|---|
| A new industry-designed benchmark shows AI still fails most real-world, long-horizon professional tasks: across 1,000+ GDP-relevant workflows, mainstream systems fully pass only 2.6% of the hardest challenges. Agents' Last Exam (ALE), mapped to O*NET occupational categories and built with 250+ experts, aims to shift evaluation toward measurable economic impact.arxiv | Hanwen Xing provider id |
2026-06-03 | 3 |
| Human-authored procedural 'Skills' lift LLM agent success by 16.2 percentage points on average—gains vary sharply by domain and sometimes harm performance—while model-generated Skills add no net value; narrowly targeted Skills let smaller models match larger ones, suggesting firms can substitute curated knowledge for compute.manual | Hanwen Xing provider id |
2026-02-13 | 150 |
Citation observation summary
Semantic Scholar supplied counts for 2 of 2 papers in this view; 0 are missing. The observed paper counts sum to 153 cumulative citations. This is a coverage summary, not an author score or h-index.