Tengyu Xu
Computer Science · The Ohio State University
Publications
53
Citations
475
Est. group size
—
Recurring co-author estimate
Active years
11
Publishing since 2016
Tengyu Xu's research centers on reinforcement learning (a machine learning approach where systems learn through trial-and-error feedback), including offline reinforcement learning, multi-agent systems, and constrained decision-making. Recent work also extends into applying reinforcement learning and optimization techniques to improve reasoning and reduce errors in large language models. Note that the publication list provided appears to mix in unrelated papers from other fields (e.g., ophthalmology, sports science), which may reflect data-matching issues rather than this researcher's actual body of work.
Publication output rose sharply around 2020-2021, dipped in 2022-2024, and picked up again in 2025, suggesting an uneven but recently reactivated publication pace.
Generated by claude-sonnet-5 from public bibliographic data · Jul 20, 2026
- A preliminary comparative study of peripheral capsular linear incision versus circular micro-tear capsulotomy for in situ lens regeneration
BMC Ophthalmology · 2026
- The efficacy of complex training for enhancing change of direction ability: A systematic review and meta-analysis
International Journal of Sports Science & Coaching · 2025
- Step-KTO: Optimizing Mathematical Reasoning through Stepwise Binary Feedback
arXiv (Cornell University) · 2025
- Think Smarter not Harder: Adaptive Reasoning with Inference Aware Optimization
arXiv (Cornell University) · 2025
- H <scp>yper</scp> Z <scp>ero</scp> : A Customized End-to-End Auto-Tuning System for Recommendation with Hourly Feedback
2025
- Learning Auxiliary Tasks Improves Reference-Free Hallucination Detection in Open-Domain Long-Form Generation
2025
- LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM Training
arXiv (Cornell University) · 2025
- Step-KTO: Optimizing Mathematical Reasoning through Stepwise Binary Feedback
2025
- Boosting LLM Reasoning via Spontaneous Self-Correction
arXiv (Cornell University) · 2025
- Learning Auxiliary Tasks Improves Reference-Free Hallucination Detection in Open-Domain Long-Form Generation
arXiv (Cornell University) · 2025
- Faster algorithm and sharper analysis for constrained Markov decision process
Operations Research Letters · 2024
- Randomized trial comparing the effects of a 3D head-up system and microscope eyepiece-assisted simulated vitrectomy with intraocular illumination on the ocular surface of an operator
BMC Ophthalmology · 2024
- Research on Fault Alarm Location Technology of Underground Cable Based on Multi-Source Information Fusion
2024
- Provably Efficient Offline Reinforcement Learning With Trajectory-Wise Reward
IEEE Transactions on Information Theory · 2024
- Comparison of the efficacy and safety of ultrasonic cycloplasty vs valve implantation and anti-VEGF for the treatment of fundus disease-related neovascular glaucoma
International Journal of Ophthalmology · 2023
- arXiv (Cornell University)×33
- BMC Ophthalmology×2
- International Journal of Heat and Mass Transfer×1
- Operations Research Letters×1
- Neural Information Processing Systems×1
- Zaiwei Chen
Computer Science · Purdue University West Lafayette
- Ziwei Guan
Computer Science · The Ohio State University
- Qinbo Bai
Computer Science · Purdue University West Lafayette
- Andrew Perrault
Computer Science · The Ohio State University
- Nan Jiang
Computer Science · Purdue University West Lafayette
This profile was generated automatically from public scholarly data (OpenAlex). Group size and activity levels are estimates derived from co-authorship patterns.
Last updated Jul 19, 2026.
Claim or correct this profile