No tags available
Rankings use category fit, feature coverage, pricing signals, public reception, and recency. Affiliate relationships do not affect scores.
AlphaZero is a computer program developed by DeepMind that utilizes generalized reinforcement learning to master board games. Released in 2017, the system learned to play chess, shogi, and Go entirely through self-play, without relying on human-derived opening books or endgame databases. Within hour...
AlphaFold 3 is an artificial intelligence model developed by Google DeepMind and Isomorphic Labs. Introduced in 2024, it predicts the joint three-dimensional structure and interactions of complexes comprising proteins, DNA, RNA, and small molecule ligands. Expanding upon the protein-focused capabili...
AlphaFold 2 is an artificial intelligence system developed by Google DeepMind to predict the three-dimensional structure of proteins from their amino acid sequences. First released in 2020, it achieved unprecedented accuracy in the 14th Critical Assessment of Protein Structure Prediction (CASP14) co...
DeepSeek-R1 is an open-weight large language model developed by the Chinese artificial intelligence company DeepSeek and released in January 2025. The model is trained using reinforcement learning techniques to enhance its chain-of-thought reasoning capabilities, specifically targeting mathematics,...
Stable Diffusion 1.5 is an open-weight text-to-image artificial intelligence model released in October 2022 by Stability AI in collaboration with Runway and LMU Munich. It utilizes a latent diffusion architecture that generates detailed images by gradually denoising a mathematical representation of...
WaveNet is a deep neural network architecture for generating raw audio waveforms, introduced by researchers at DeepMind in a paper published in September 2016. It uses a dilated causal convolutional network to model audio signals directly at the sample level, producing more natural-sounding speech t...
Whisper is an automatic speech recognition model released by OpenAI in September 2022 as open-source software. It was trained on approximately 680,000 hours of multilingual and multitask supervised data collected from the web, covering numerous languages and dialects. The model is designed to perfor...
SAM (Segment Anything Model) is a vision foundation model developed by Meta AI and released in 2023. It was trained on the SA-1B dataset of over one billion segmentation masks on eleven million images, enabling zero-shot generalization to segment objects not present in training data. Users can promp...
MuZero is a reinforcement learning algorithm developed by DeepMind, detailed in a 2020 publication in the journal Nature. It learns to master environments without being provided their rules by simultaneously learning a model of the environment and improving its decision-making policy. MuZero achieve...
CLIP (Contrastive Language–Image Pretraining) is a neural network model introduced by OpenAI in 2021. It is trained on approximately four hundred million image and text pairs collected from the internet using a contrastive objective that aligns image and text representations in a shared embedding sp...
DeepSeek-V3 is a large language model released by the Chinese AI company DeepSeek in late 2024. It is built as a Mixture-of-Experts (MoE) model with a total of 671 billion parameters, of which only 37 billion are activated per token. The model was trained efficiently using a specialized architecture...
Chinchilla is a large language model developed by DeepMind and detailed in a 2022 research paper. It features 70 billion parameters and was trained using the same computational budget as DeepMind's earlier 280-billion-parameter Gopher model. By adjusting the balance between model size and training d...
Flux.1 Pro is a text-to-image artificial intelligence model developed by Black Forest Labs and released in August 2024. The model is positioned as the flagship tier of the Flux.1 family, offering high prompt adherence and image detail generation. Unlike the company's open-weights versions, such as t...
Claude 3.7 Sonnet is a large language model released by the AI research company Anthropic in 2025. It operates as a hybrid system, allowing users to toggle between a standard fast-response mode and an extended "thinking" mode for complex problem-solving. The extended thinking mechanism enables the m...
Tacotron 2 is a neural network architecture for text-to-speech (TTS) synthesis introduced by Google researchers in 2017. The system operates by splitting the process into two stages: a sequence-to-sequence model predicts mel-scale spectrograms from input text, and a modified WaveNet model acts as a...
Llama 3.1 405B is a large language model released by Meta in 2024, serving as the flagship of the Llama 3.1 collection. It features 405 billion parameters and supports a context window of 128,000 tokens. As an open-weight model, it was made available for download, providing developers with a tool co...
Flux.1 Dev is a text-to-image diffusion model released in 2024 by Black Forest Labs, the company founded by former members of Stability AI. It is a guidance-distilled variant of the Flux.1 Pro model, released under a non-commercial license for research and development. The model uses a flow-matching...
Midjourney v4 is a version of the Midjourney text-to-image artificial intelligence model released in 2022. It represented a significant architectural update trained on a new codebase and dataset, resulting in improved image coherence, higher resolution options, and better handling of complex, multi-...
SAM 2 (Segment Anything Model 2) is an artificial intelligence model developed by Meta and released in 2024. It extends the capabilities of the original Segment Anything Model from static images to video, utilizing a memory mechanism to track and segment objects across frames in real time. The model...
DINOv2 is a self-supervised vision foundation model developed by Meta AI and released in 2023. It was trained on a highly curated dataset of 142 million images without relying on manual labels or text supervision. By utilizing an improved student-teacher architecture, the model produces robust visua...
Sora is a text-to-video generative artificial intelligence model developed by OpenAI and announced in February 2024. The model utilizes a diffusion transformer architecture to synthesize high-definition video clips from natural language text prompts. It is capable of generating up to one minute of f...
Grok-3 is xAI's third-generation large language model, released in 2025. Trained on a massive computing cluster, it represents xAI's continued effort to compete with frontier models from OpenAI and Anthropic. Grok models are designed to integrate with the X platform and are marketed as having fewer...
GPT-3 is a large language model developed by OpenAI and released in June 2020. With 175 billion parameters, it represented a significant scale-up from previous models and helped establish few-shot learning as a powerful paradigm for natural language processing. GPT-3 demonstrated capabilities in tex...
Released in 2024 by the AI firm DeepSeek, DeepSeek Coder V2 is an open-weight, mixture-of-experts large language model designed for code generation and software engineering tasks. The model features a massive parameter count with an activated subset of experts for each query, enabling it to process...
InstructGPT is a family of large language models introduced by OpenAI in 2022, designed to align artificial intelligence outputs with human intent. Developed as an evolution of the GPT-3 model, it was trained using reinforcement learning from human feedback (RLHF) to follow specific instructions rat...
Mamba is a deep learning architecture introduced in 2023 by researchers Albert Gu and Tri Dao that utilizes selective state space models (SSMs) for natural language processing. Unlike traditional Transformer models that require quadratic computational complexity for sequence length, Mamba achieves l...
Qwen2.5 is an open-weight family of large language models released in 2024 by Alibaba Cloud. The series includes base and instruction-tuned models ranging in size from 0.5 billion to 72 billion parameters. Notable for its strong multilingual capabilities and coding proficiency, the architecture is d...
You're in. We'll email you when new Model entries land.