Direction, evidence grade, and study type are AI-generated labels (gpt-5-mini), not human-verified. Syntheses are LLM-written. "Tensions" are machine-detected candidates, not confirmed contradictions. A research-acceleration tool, not peer review.
How this is built →
1Distinct papers
8Unique collaborators
1/1OpenAlex citation coverage
Publication span: 2025. Corpus fetch span: 2026.
Identity provenance
Provider IDs
No provider ID is stored.
ORCID evidence
No valid ORCID is stored.
Observed aliases (1)
- Dong, Xuanyu (openalex, provider refresh)
Topics and outcomes in this view
Assessment themes
- Adoption: 1 paper
- Human Ai Collab: 1 paper
- Productivity: 1 paper
Claim outcomes
- Other: 1 paper
- Output Quality: 1 paper
- Task Completion Time: 1 paper
Papers in the OpenAlex view
Latest stored OpenAlex author observations only. Citation counts below are from the same provider and are not combined with other services.
Scroll the table horizontally to see every column.
| Paper | Author evidence | Date | Provider citations |
|---|---|---|---|
| A new benchmark of real-world finance workflows finds top AI agents complete fewer than 40% of tasks: GPT-5.1 spends nearly 17 minutes per workflow yet passes only 38.4% of cases, exposing persistent failure modes on messy, multimodal enterprise work.openalex | Dong, Xuanyu unresolved |
2025-12-15 | 0 |
Citation observation summary
OpenAlex supplied counts for 1 of 1 papers in this view; 0 are missing. The observed paper counts sum to 0 cumulative citations. This is a coverage summary, not an author score or h-index.