AI Models
DistilGPT2 is a distilled, English-language text generation model derived from GPT-2. It offers a faster and lighter alternative to its predecessor, making it suitable for various…
Discover the best AI Models — from language transformers like DistilGPT2 and Claude Opus to advanced systems such as GPT-5.6: Frontier intelligence and Gemini 3.1 Pro by Google DeepMind, this is where developers and researchers explore tools that enhance their projects: generating human-like text, building conversational agents, developing complex algorithms, and pushing the boundaries of AI capabilities. Browse 324 AI Models spanning AI Models & LLMs, Other AI Tools, Machine Learning Platforms, Scientific Discovery & Lab Tools and AI Image Generators and more.
324 tools
AI Models
DistilGPT2 is a distilled, English-language text generation model derived from GPT-2. It offers a faster and lighter alternative to its predecessor, making it suitable for various…
AI Models
Claude Opus is a hybrid reasoning model designed for serious coding and AI agents, featuring a 1M context window. It excels in production-ready code, sophisticated AI agents, and…
GPT-5.6 is OpenAI's latest AI model, designed to enhance productivity and efficiency across various tasks. It offers improved performance per dollar, enabling users to achieve…
AI Models
Gemini 3.1 Pro is an advanced AI model designed for complex tasks and deep reasoning. It excels in multimodal understanding, providing smart and concise responses, making it ideal…
AI Models
Grok 4.6 enhances the capabilities of Grok 4.5, focusing on long-running agents and ambitious interactive and visual tasks. It excels in complex projects, offering improved…
AI Models
Stability AI offers a suite of open-source generative AI models for creating images, video, audio, and 3D content. These innovative solutions cater to creators and enterprises…
AI Models
Qwen Studio provides a versatile AI model that excels in chatbot functionality, image and video understanding, image generation, document processing, and web search integration,…
GLM-5.3 is an advanced AI model designed for coding and cyber capabilities. It offers significant improvements in complex coding tasks and vulnerability discovery, making it a…
AI Models
Cohere Command is a family of scalable AI language models designed for enterprise use. It offers high performance and accuracy for real-world agentic applications, enabling…
AI Models
Nano Banana Pro is an AI model designed for creating and editing images with studio-quality precision and control. It leverages advanced artificial intelligence to generate…
AI Models
GPT Image 2 is an advanced image generation model by OpenAI, designed for fast and high-quality image creation and editing. It supports various image sizes and high-fidelity…
FLUX.2 is the next generation of image generation from Black Forest Labs, offering state-of-the-art quality, speed, and controllability for AI-generated images, making it ideal…
AI Models
Muse Glimmer is a 30-billion-parameter causal language model designed for autonomous agentic tasks on consumer hardware. It integrates multi-step reasoning, reliable tool use, and…
AI Models
Midjourney V8.2 is an AI model that allows users to explore and switch between different versions using the version parameter. It focuses on aesthetics, image quality, and…
Kling 3.0 is an advanced AI video and image generator that transforms text, images, and references into high-quality multimodal content. It offers tools for cinematic narratives…
Runway Gen-4.5 is an advanced AI video generation model that enhances motion quality, prompt adherence, and temporal consistency. It is available on all paid plans, making it a…
AI Models
AI21 builds enterprise-grade Foundation Models and AI Systems, focusing on accuracy, reliability, and scalability. Their solutions power critical business workflows, offering…
AI Models
Gemma is a collection of lightweight, open models built from the same technology that powers Gemini models. It enables developers to create AI applications for various platforms,…
AI Models
Wan is an AI creative platform designed to lower the barriers to creative work. It offers features like text-to-image, image-to-image, text-to-video, image-to-video, and image…
AI Models
Mistral Large 3 is a state-of-the-art AI model designed for enterprises, enabling customization, fine-tuning, and deployment of AI assistants and agents. It features a sparse…
Eleven v3 is an advanced AI voice model that generates lifelike speech in over 70 languages. It offers emotional depth, multi-speaker control, and is designed for various…
Suno v5.5 is an AI music model that enhances creativity with features like Voices, Custom models, and My Taste, allowing users to express their musical identity. It caters to both…
AI Models
Genie 3 is a groundbreaking AI model that generates photorealistic environments from simple text descriptions. It allows users to explore these interactive worlds in real-time,…
AI Models
NVIDIA Nemotron 3 Ultra is a powerful AI model designed for complex reasoning and multilingual tasks. With 550 billion parameters, it excels in long-context analysis and tool use,…
AI Models
gpt-oss includes two open-weight language models, gpt-oss-120b and gpt-oss-20b, designed for strong performance and efficient deployment. They are available under the Apache 2.0…
AI Models
Muse Spark 1.2 is an AI model optimized for real coding workflows, offering higher first-attempt accuracy and reliable tool calling. It is designed to enhance coding capabilities…
Amazon Nova is a suite of foundation models and services designed for enterprise-scale applications, offering high performance and multimodal capabilities. It enables…
AI Models
MiniMax-M2.5 is an advanced AI model designed for coding, office work, and agentic tool use. It excels in efficiency and cost-effectiveness, enabling innovative applications…
AI Models
Mistral Small 4 is a versatile AI model that unifies reasoning, coding, and multimodal capabilities into a single platform. It allows users to customize, fine-tune, and deploy AI…
AI Models
Command A+ is a Mixture of Experts model with 25B active and 218B total parameters, designed for complex reasoning, vision, and multilingual tasks across 48 languages, providing…
AI Models
ERNIE 5.1 is a cutting-edge AI model that excels in performance while minimizing pre-training costs. It offers significant advancements in agent capabilities, reasoning, and…
AI Models
Gemini Omni is an AI model from Google DeepMind that allows users to create and edit videos through natural conversation. It combines reasoning with creative capabilities,…
AI Models
Imagen is a cutting-edge text-to-image AI model developed by Google DeepMind. It generates photorealistic images with exceptional clarity and speed, allowing users to bring their…
AI Models
Qwen-Image is an advanced image generation foundation model that excels in complex text rendering and precise image editing. It supports a variety of artistic styles and offers…
AI Models
Ideogram 4.0 is an advanced AI model designed for various applications. It offers robust capabilities for users seeking to leverage AI technology for their projects, enhancing…
AI Models
Stable Diffusion 3.5 is a Multimodal Diffusion Transformer model designed for text-to-image generation. It enhances image quality, typography, and prompt understanding while being…
AI Models
Ray3.2 is a versatile AI video model that transforms creative intent into scalable video workflows. It offers enhanced control, continuity, and cinematic direction across multiple…
AI Models
LTX-2 is a DiT-based audio-video foundation model designed to generate synchronized video and audio. It combines modern video generation techniques with open weights, enabling…
MiniMax Hailuo 2.3 is an advanced AI model designed for complex video performance and media tasks. It aims to enhance user experience with its multi-modal capabilities, making it…
AI Models
Cosmos 3 is an advanced AI model developed by NVIDIA, designed to enhance various applications in artificial intelligence, high-performance computing, and robotics. It leverages…
AI Models
Marble is a platform that allows users to create and share immersive 3D worlds. It provides tools for building detailed environments, making it ideal for creators and developers…
AI Models
AlphaFold 3 is an AI model developed by Google DeepMind that predicts protein structures and interactions, significantly accelerating biological research and understanding of…
AI Models
SAM 3 allows users to utilize text and visual prompts to accurately identify, segment, and track objects in images or videos. It will soon be available in Instagram Edits and…
AI Models
Lyria 3.5 is an advanced music generation model by Google DeepMind that helps users compose songs with ease. It offers technical control and the ability to create tracks in…
Sonic is Cartesia's real-time text-to-speech API that generates expressive voices with laughter in 44 languages. Designed for AI agents and interactive applications, it offers low…
AI Models
Sesame CSM is a conversational speech model that generates audio codes from text and audio inputs. It utilizes a Llama backbone and is designed for research and educational…
AI Models
IBM Granite is a family of open, trusted AI models for business, offering efficient language, vision, speech, and guardrail models that can be customized and deployed to meet…
Inkling is a 975B-parameter multimodal AI model designed for developers. It accepts text, image, and audio inputs, generating text outputs for various applications, including…
Vidu Q3 is an AI video generation model that creates videos with native audio, enabling seamless storytelling. It offers faster motion quality and stronger prompt control, making…
HunyuanImage 3.0 is a powerful native multimodal model designed for image generation. It excels in both text-to-image and image-to-image tasks, offering advanced capabilities for…
AI Models
Olmo from Ai2 is a fully open language model designed for advanced AI research and applications. It offers various model variants optimized for programming, reasoning, and…
Phi-4-reasoning-vision-15B is a multimodal AI model developed by Microsoft, designed for tasks requiring vision-language understanding and reasoning capabilities. It excels in…
ESMFold2 predicts high-resolution, all-atom 3D protein structures from amino acid sequences. It offers open access to both single-sequence and MSA-conditioned models, enhancing…
AI Models
NVIDIA Isaac GR00T N1.7 is an open foundation model designed for humanoid robot reasoning and skills. It processes multimodal inputs to perform manipulation tasks, allowing…
AI Models
Falcon-H1R-7B is a reasoning-specialized AI model designed to enhance performance in mathematics, programming, and logic tasks. Developed by the Technology Innovation Institute,…
AI Models
OpenAI is a leading artificial intelligence research and deployment company. It focuses on developing advanced AI models and making them accessible for various applications. The…
AI Models
AudioGen is an auto-regressive generative AI model that creates audio samples based on descriptive text captions. It addresses challenges in audio generation, such as separating…
AI Models
OPT-175B is a 175-billion-parameter language model released by Meta AI for the research community. It provides access to large-scale language models, enabling deeper understanding…
AI Models
Gato is a single generalist AI agent developed by Google DeepMind. It can perform a wide variety of tasks, including playing Atari games, captioning images, chatting, and…
AI Models
NeRF, or Neural Radiance Fields, synthesizes novel views of complex scenes by optimizing a continuous volumetric scene function. It uses a sparse set of input views to represent…
AI Models
DETR is an end-to-end object detection and panoptic segmentation framework that integrates Transformers as a core component. It simplifies the architecture, directly predicts…
AI Models
Engineering at Meta is the official blog for Meta's engineering team, sharing insights into their work. It covers AI, data infrastructure, development tools, and more. The site…
AI Models
Densenet is a deep convolutional neural network architecture available through PyTorch. It enhances feature propagation and reuse by connecting each layer to every other layer in…
MobileNets are a family of mobile-first computer vision models for TensorFlow, designed for efficient on-device or embedded applications. They maximize accuracy while minimizing…
AI Models
SqueezeNet is a deep convolutional neural network model available through PyTorch. It achieves AlexNet-level accuracy with significantly fewer parameters and a smaller model size,…
AI Models
Xception is a deep learning model available through the Keras 3 API. It is part of Keras Applications, offering pre-trained models for various computer vision tasks. Xception is…
AI Models
Explore the inner workings of artificial neural networks with Inceptionism. This technique visualizes network layers by enhancing input images to reveal learned features, aiding…
AI Models
NVIDIA leads in AI computing, driving advancements in GPUs, AI infrastructure, and agentic AI. They offer solutions for frontier model training, data center AI factories, and…
AI Models
Transfo-XL-WT103 is a causal transformer model with relative positioning embeddings that can reuse hidden states for longer context. Developed by Zihang Dai and others, it uses…
AI Models
MusicLM is an AI model that generates high-fidelity music from text descriptions. It can produce music up to 24 kHz that remains consistent over several minutes, outperforming…
AI Models
Jukebox is a neural network that generates music, including rudimentary singing, as raw audio. It can produce music in various genres and artist styles, offering a novel approach…
AI Models
Google DeepMind explores large language models like Gopher, focusing on their capabilities, ethical considerations, and efficient training. Research includes a 280 billion…
AI Models
This research empirically analyzes the optimal trade-off between model size and training data for large language models given a fixed compute budget. It reveals that current large…
AI Models
OpenAI Codex is an AI system that translates natural language into code. It interprets simple commands and executes them, making it possible to build natural language interfaces…
AI Models
AlphaCode is an AI system developed by Google DeepMind that writes computer programs at a competitive level. It achieved an estimated rank within the top 54% of participants in…
Parti is an autoregressive text-to-image generation model that creates high-fidelity photorealistic images. It treats image generation as a sequence-to-sequence problem,…
AI Models
LaMDA is Google's breakthrough conversational AI model, designed to engage in free-flowing dialogue across a vast array of topics. It builds upon Transformer architecture, trained…
AI Models
The Google AI Blog Archive serves as a historical repository for posts related to artificial intelligence research and developments from Google. It allows users to navigate…
AI Models
InstructGPT models are AI language models trained to better follow user intentions than GPT-3. They are more truthful, less toxic, and aligned with user goals through…
AI Models
DreamFusion is an AI model that generates 3D objects from text descriptions. It leverages pre-trained 2D diffusion models to create relightable 3D assets like Neural Radiance…
AI Models
Point E is an AI model developed by OpenAI, hosted on Hugging Face Spaces. It focuses on generating 3D point clouds from text prompts. This tool allows users to explore AI-driven…
AI Models
Diffuse The Rest is a Hugging Face Space that generates images from text descriptions using a Stable Diffusion model. Simply type what you want to see, and the AI creates a…
AI Models
Graph Attention Networks (GATs) are novel neural network architectures designed for graph-structured data. They leverage masked self-attentional layers to address limitations of…
AI Models
Graph Convolutional Networks (GCNs) are a type of neural network designed to process data structured as graphs. They generalize convolutional neural networks to graph-structured…
AI Models
FaceNet is a TensorFlow implementation for face recognition and clustering, based on the FaceNet paper. It leverages deep learning models to generate unified embeddings for faces,…
AI Models
Longformer Base 4096 is a transformer model designed for processing long documents. It builds upon RoBERTa, pretrained on extended sequences up to 4,096 tokens. This model employs…
AI Models
The BigBird base model is a transformer that extends BERT to handle much longer sequences using block sparse attention. It is pre-trained on English text for masked language…
AI Models
The NVIDIA Technical Blog offers insights into AI, HPC, and accelerated computing. It features articles on cutting-edge research, developer tools, and industry trends. This…
AI Models
RAG (Retrieval-Augmented Generation) combines pretrained language models with external data sources. It fetches relevant passages to condition generation, enhancing factual…
AI Models
SimCLR is a framework for self-supervised and semi-supervised learning of visual representations. It simplifies previous approaches by using contrastive learning to maximize…
AI Models
DeepMind is a leading artificial intelligence research laboratory focused on advancing the state of the art in AI. They develop cutting-edge AI models and systems, pushing the…
AI Models
Phi-2 is a 2.7 billion-parameter language model from Microsoft Research. It demonstrates outstanding reasoning and language understanding, achieving state-of-the-art performance…
AI Models
ControlNet enhances diffusion models by adding conditional control, allowing users to guide image generation with specific inputs. It enables fine-tuning without destroying…
AI Models
This AI model details a large, deep convolutional neural network trained for ImageNet classification. It achieved state-of-the-art results with top-1 and top-5 error rates of…
AI Models
This GitHub repository provides an example implementation of Deep Convolutional Generative Adversarial Networks (DCGAN) using PyTorch. It allows users to train models on datasets…
AI Models
This article provides an in-depth explanation of the Wasserstein GAN paper, making its complex theory more accessible. It details the paper's importance, theoretical…
AI Models
This research introduces a novel training methodology for Generative Adversarial Networks (GANs). It progressively grows both the generator and discriminator, starting from low…
AI Models
StyleGAN3 introduces alias-free generative adversarial networks by overhauling signal processing within the generator. This innovation eliminates "texture sticking," enabling more…
AI Models
Google DeepMind's publication archive showcases cutting-edge AI research. Explore recent studies on complex challenges, from AI consciousness and safety to advanced vision and…
AI Models
GitHub Pages is a static site hosting service offered by GitHub. It allows users to publish websites directly from a GitHub repository. Ideal for project documentation, personal…
This AI model research paper explores the approximation and convergence properties of Generative Adversarial Networks (GANs). It addresses fundamental questions about how…
AI Models
Graphormer is a deep learning package for molecule modeling tasks, accelerating research in material and drug discovery. It supports molecular dynamics and property prediction,…
AI Models
reCAPTCHA is a security measure designed to protect websites from automated bots and malicious traffic. It verifies that users are human before granting access to web pages,…
AI Models
UL2 20B is an open-source unified language learner model that unifies various language modeling paradigms. It improves performance across fine-tuning and few-shot learning tasks…
AI Models
FLAN-T5 is an enhanced version of the T5 language model, specifically fine-tuned on a diverse mixture of tasks. It offers improved performance without requiring further…
AI Models
PaLM-E is an embodied multimodal language model that integrates real-world continuous sensor data with text. It enables robots to perform complex tasks by grounding language in…
This framework introduces a novel active learning method for enriching large 3D shape datasets with semantic region annotations. It efficiently combines manual annotation,…
AI Models
DGCNN is an AI model for learning on point clouds, implementing Dynamic Graph CNN. It achieves state-of-the-art performance in tasks like classification and segmentation of 3D…
AI Models
MMAction2 is a foundational library for action recognition and video understanding tasks. It provides a comprehensive toolkit for developing and deploying state-of-the-art video…
AI Models
Wav2vec 2.0 is a self-supervised learning algorithm for automatic speech recognition. It learns from raw audio, requiring minimal transcribed data to achieve high accuracy. This…
AI Models
AudioLM is an AI model that generates high-quality audio with long-term consistency. It treats audio generation as a language modeling task, mapping audio to discrete tokens. The…
AI Models
Jiasen Lu is a Research Scientist at the Allen Institute for AI. His work focuses on advancing artificial intelligence research. This page serves as a personal homepage, providing…
AI Models
This listing refers to a GitHub Pages site that is not found. It indicates that there is no active GitHub Pages site at the provided domain. Users attempting to publish a site are…
ALIGN is an AI model that scales visual and vision-language representation learning using noisy text supervision from over one billion image-alt-text pairs. It achieves…
AI Models
DALL·E is an AI model that generates images from text descriptions. It can create a wide range of visual concepts, combine unrelated ideas, render text, and apply transformations…
AI Models
Flamingo is a single visual language model (VLM) from Google DeepMind that excels at few-shot learning across diverse multimodal tasks. It processes interleaved images, videos,…
AI Models
The Google AI Blog is a central hub for announcements, research updates, and insights into Google's advancements in artificial intelligence. It covers a wide range of AI topics,…
AI Models
BEiT is a self-supervised vision representation model that uses masked image modeling to pre-train vision transformers. It tokenizes images into visual tokens and recovers masked…
AI Models
The ACM Digital Library is a comprehensive research portal providing access to a vast collection of scholarly articles, conference proceedings, and journals in computing. It…
AI Models
SchNetPack is a Python library for deep learning for atomistic systems. It provides a flexible framework for building and training neural network potentials, enabling researchers…
AI Models
Variational autoencoders (VAEs) offer a probabilistic approach to latent space representation. Unlike standard autoencoders, VAEs encode observations into probability…
AI Models
Meta AI is a research initiative focused on advancing artificial intelligence through open research and accessible tooling. It aims to connect people with what they care about and…
AI Models
Reformer is an AI model that enhances the Transformer architecture for processing extensive sequential data. It addresses limitations in attention mechanisms and memory…
AI Models
AlphaDev is an AI system that uses reinforcement learning to discover enhanced computer science algorithms. It has successfully identified faster sorting and hashing algorithms,…
AI Models
ImageBind is a multimodal AI model from Meta AI that binds data from six modalities: image, video, audio, text, depth, and thermal. It learns a single embedding space without…
AI Models
Intern Large Models, developed by Shanghai AI Laboratory, offers open-source LLMs and MLLMs. Their suite includes models like InternVL and InternLM-XComposer, alongside a…
AI Models
通义实验室 (Meet Tongyi) is the official website for Alibaba Cloud's Tongyi Qianwen large language models. It showcases the full suite of models, the latest industry news, and…
AI Models
Zephyr-7B-β is an open-source language model fine-tuned from Mistral-7B-v0.1 using Direct Preference Optimization. It excels as a helpful assistant, demonstrating strong…
AI Models
VideoPoet is a large language model from Google Research capable of zero-shot video generation. It transforms autoregressive language models into high-quality video generators,…
AI Models
Lumiere is a space-time diffusion model from Google Research for generating realistic, diverse, and coherent videos. It synthesizes entire video durations in a single pass,…
AI Models
Fuyu-8B is an open-source multimodal AI model designed for digital agents. Its simplified architecture supports arbitrary image resolutions, enabling it to answer questions about…
AI Models
IDEFICS is an open-access visual language model that reproduces state-of-the-art capabilities. It accepts interleaved image and text inputs to generate text outputs, comparable to…
AI Models
MiniGPT-4 is an AI model that enhances vision-language understanding by aligning a frozen visual encoder with a large language model. It can generate detailed image descriptions,…
AI Models
LLaVA is a large multimodal model that combines a vision encoder with a language model for general-purpose visual and language understanding. It excels at multimodal chat…
AI Models
Mamba is a novel state space model architecture designed for efficient sequence modeling, particularly effective on information-dense data like language. It offers a…
AI Models
RWKV is a powerful language model that offers efficient inference and flexible fine-tuning capabilities. It supports various applications, including desktop GUIs, web-based…
AI Models
OpenELM is a state-of-the-art open language model family from Apple Machine Learning Research. It features a layer-wise scaling strategy for enhanced accuracy and provides a…
AI Models
The Institute for Machine Learning at Johannes Kepler University Linz conducts renowned research and provides education in machine learning. It focuses on developing and applying…
AI Models
Medium is a vast online publishing platform where individuals and organizations share diverse articles, stories, and insights. It hosts content on a wide array of topics, from…
AI Models
N-BEATS is a neural-network based model for univariate time-series forecasting. Developed by ServiceNow Research, it implements the N-BEATS algorithm for reproducible experimental…
AI Models
零一万物 (01.AI) is a global AI company focused on AI 2.0, driven by foundation model breakthroughs. They aim to revolutionize technology, platforms, and applications, creating a new…
AI Models
Physics Informed Neural Networks (PINNs) solve supervised learning tasks by respecting physical laws described by nonlinear partial differential equations. They enable data-driven…
AI Models
Ansuman Bora's personal portfolio website showcases his professional background as a Software Engineer at Corteva Agriscience. It highlights his academic achievements, including a…
AI Models
Octo is an open-source, generalist robot policy designed for broad applicability in robotic manipulation. This transformer-based diffusion policy is pretrained on a large dataset,…
AI Models
APS Journals is a digital platform providing access to a vast collection of peer-reviewed scientific research articles published by the American Physical Society. It serves as a…
AI Models
DBRX is a state-of-the-art open large language model from Databricks, excelling in benchmarks for language, programming, and math. It offers improved efficiency and quality,…
AI Models
Fast R-CNN is a deep learning framework for object detection. It significantly speeds up training and testing compared to previous methods like R-CNN and SPPnet, achieving higher…
AI Models
py-faster-rcnn is a Python implementation of the Faster R-CNN object detection model. It offers an alternative to the official MATLAB version, providing similar accuracy with…
AI Models
Caffe framework with SSD implementation for object detection. This repository provides a fast, open framework for deep learning, specifically tailored for the Single Shot MultiBox…
AI Models
Feature Pyramid Networks (FPN) enhance object detection by creating multi-scale feature maps from deep convolutional networks with minimal computational overhead. This…
ToolPotion is the AI tools directory people browse to find the best AI Models. Listing is free and every submission is editor-reviewed — put your tool where the search starts.