GPT-4 Technical Report
Explore this paper's citation graph
Summary
GPT-4, a large-scale, multimodal model which can accept image and text inputs and produce text outputs, is developed, a Transformer-based model pre-trained to predict the next token in a document which exhibits human-level performance on various professional and academic benchmarks.
- Published
- 2023-03-15
- Cited by
- 26,898
- References
- 0
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:257532815
References
No references recorded for this paper.
Cited by
- Transfer Learning in Deep Reinforcement Learning: A Survey
- Model-Based Deep Learning
- Data extraction methods for systematic review (semi)automation: Update of a living systematic review
- Mat2Stencil: A Modular Matrix-Based DSL for Explicit and Implicit Matrix-Free PDE Solvers on Structured Grid
- Engineering flexible machine learning systems by traversing functionally invariant paths
- Synopses of Movie Narratives: a Video-Language Dataset for Story Understanding
- Worldwide AI ethics: A review of 200 guidelines and recommendations for AI governance
- What is Flagged in Uncertainty Quantification? Latent Density Models for Uncertainty Categorization
- KG2Tables: A Domain-Specific Tabular Data Generator to Evaluate Semantic Table Interpretation Systems (Resource Paper)
- From Natural Language to Simulations: Applying GPT-3 Codex to Automate Simulation Modeling of Logistics Systems
- The alignment problem from a deep learning perspective
- Quantitative AI Risk Assessments: Opportunities and Challenges
- Do Large Language Models know what humans know?
- Training Overparametrized Neural Networks in Sublinear Time
- PoliGraph: Automated Privacy Policy Analysis using Knowledge Graphs
- Neural Theory-of-Mind? On the Limits of Social Intelligence in Large LMs
- Broken Neural Scaling Laws
- Refiner: Data Refining against Gradient Leakage Attacks in Federated Learning
- Human-like intuitive behavior and reasoning biases emerged in large language models but disappeared in ChatGPT
- Elixir: Train a Large Language Model on a Small GPU Cluster
Related papers
No related papers recorded.