TL;DR
Andrej Karpathy has announced the release of Pelican, an AI model designed for advanced image and language processing. The development is confirmed, but its full capabilities and impact are still still emerging.
OpenAI researcher Andrej Karpathy has publicly announced the release of Pelican, a new AI model focused on advanced image and language processing capabilities. This development is confirmed by Karpathy’s official status update and marks a notable milestone in AI research efforts.
According to Karpathy’s official post on X (Twitter), Pelican is designed to push the boundaries of current AI models, particularly in multi-modal tasks involving both images and text. The specifics of Pelican’s architecture, training data, and deployment are not yet fully disclosed, but initial indications suggest it aims to outperform existing models in certain benchmarks.
Karpathy emphasized that Pelican is still in early stages of testing, with further updates expected as development progresses. Learn more about related aerospace and defense AI benchmarks in this article. The model’s initial deployment appears to be targeted toward research environments, with potential future applications in industry and academia.
Implications of Pelican for AI Innovation
The release of Pelican signifies a step forward in the development of multi-modal AI systems, which combine visual and textual understanding. This could lead to more sophisticated AI applications in areas such as autonomous systems, content creation, and data analysis. For researchers and industry stakeholders, Pelican’s capabilities may influence future AI benchmarks and standards.

AI Image Generation with Prompt Engineering: Create Stunning Visuals with AI Tools (AI Prompting Secrets: Unlocking Creativity, Automation, and Efficiency)
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Background on Karpathy’s AI Research and Pelican’s Development
Andrej Karpathy, a prominent figure in AI research and former Tesla AI director, has a history of contributing to deep learning advancements. His recent activities have focused on multi-modal models that integrate vision and language, exemplified by projects like OpenAI’s CLIP and DALL·E.
Pelican appears to be a continuation of this research trajectory, aiming to improve upon existing models’ ability to understand and generate complex visual and textual data. The announcement comes amid ongoing industry efforts to develop more versatile and capable AI systems.
“Pelican is an early-stage model designed to explore new frontiers in multi-modal AI. We’re excited about its potential to advance research and practical applications.”
— Andrej Karpathy

Cloud-based Multi-Modal Information Analytics: A Hands-on Approach (Chapman & Hall/CRC Cloud Computing for Society 5.0)
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Unclear Aspects of Pelican’s Capabilities and Deployment
Details about Pelican’s specific architecture, training datasets, and performance metrics remain undisclosed. It is also unclear how widely Pelican will be adopted outside of research contexts or what its commercial applications might be in the near term. Further updates from Karpathy or associated teams are awaited to clarify these points.

Fine-Tuning Large Language Models: From Custom Datasets to High-Performance AI Models Using Modern Toolchains
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Next Steps in Pelican’s Development and Evaluation
Karpathy and his team are expected to publish more detailed technical information and performance evaluations of Pelican in upcoming research papers or presentations. Monitoring these developments will be key to understanding Pelican’s impact and potential integration into broader AI ecosystems.

Delphi in all its glory: AI-assisted development: Tested on 150000 lines of code: An honest guide to AI for the Delphi programmers
As an affiliate, we earn on qualifying purchases.
As an affiliate, we earn on qualifying purchases.
Key Questions
What is Pelican designed to do?
Pelican is an AI model focused on multi-modal tasks involving both image and language understanding, aiming to improve capabilities in these areas.
Who is Andrej Karpathy?
Andrej Karpathy is a prominent AI researcher, known for his work on deep learning and former director of AI at Tesla. He is now involved in advancing multi-modal AI models.
When was Pelican announced?
Pelican was announced publicly in March 2024 via Karpathy’s official status update on X (Twitter).
Will Pelican be available for commercial use?
It is not yet clear whether Pelican will be commercially available or primarily used for research; further updates are expected.
What makes Pelican different from existing models?
Pelican aims to push the boundaries of multi-modal AI, potentially offering improved performance in understanding and generating visual and textual data, though specific technical differences are not yet disclosed.
Source: hn