LLMs on University-Level Physics Coding
W. Yeadon, Alex Peach, Craig P. Testrow
Evaluates ChatGPT variants on university-level physics coding assignments, finding students outperform AI and human evaluators detect AI work with 85.3% accuracy.
A comprehensive index of artificial intelligence and machine-learning research with AI-generated summaries, citation metrics, and direct links to papers and code.
W. Yeadon, Alex Peach, Craig P. Testrow
Evaluates ChatGPT variants on university-level physics coding assignments, finding students outperform AI and human evaluators detect AI work with 85.3% accuracy.
Michael Dowling, Brian M. Lucey
ChatGPT can assist finance research by generating ideas and identifying data, but is weaker on literature synthesis and testing frameworks, with output quality depending on private data and domain expertise.
Steffen Herbold, Annette Hautli-Janisz, Ute Heuer, et al.
A large-scale study comparing human-written and ChatGPT-generated argumentative essays finds AI essays rated higher in quality by teachers, with distinct linguistic characteristics.
Katharina Jeblick, Balthasar Schachtner, Jakob Dexl, et al.
This exploratory case study evaluates the quality of simplified radiology reports generated by ChatGPT, finding that while most radiologists rated them factually correct and complete, instances of errors and potential harm remain.
Maanak Gupta, Charankumar Akiri, Kshitiz Aryal, et al.
This paper explores the dual-use of generative AI in cybersecurity, demonstrating attack techniques like jailbreaks and prompt injection, while also proposing defensive applications.
Humaid Al Naqbi, Zied Bahroun, Vian Ahmed
A PRISMA-based literature review of 159 papers analyzing how generative AI enhances productivity across multiple sectors, with bibliometric identification of ChatGPT as a dominant tool.
Jürgen Rudolph, Shannon Tan, Samson Tan
This paper compares major chatbots (ChatGPT, Bing Chat, Bard, Ernie) for higher education, finding no A-grade performers despite hype, with GPT-4 leading but Bing Chat and Bard failing.
Sai Vemprala, Rogerio Bonatti, Arthur Bucker, et al.
This paper presents a strategy combining prompt engineering and a high-level function library to enable ChatGPT for diverse robotics tasks, and introduces the PromptCraft research tool.
Inês Carvalho, Stanislav Ivanov
This paper outlines the applications, benefits, and risks of ChatGPT and large language models in tourism, establishing a research agenda for their implications.
Unknown
CriticGPT uses RLHF to train a GPT-4-based model that critiques ChatGPT code outputs, helping humans catch bugs more accurately.
Juan S Izquierdo-Condoy, Marlon Arias-Intriago, Andrea Tello-De-la-Torre, et al.
This viewpoint paper critically assesses how generative AI tools like ChatGPT affect critical thinking and cognitive autonomy in medical education, finding both potential benefits and risks of overreliance.
Unknown
Orca trains a 13B LLM to imitate GPT-4 reasoning via explanation tuning, using ChatGPT as a teacher assistant for progressive learning.