Thursday, July 16, 2026
- Interim Measures for the Management of Generative AI Servicesv4The Interim Measures for the Management of Generative Artificial Intelligence Services (Chinese: 生成式人工智能服务管理暂行办法, Shengchengshi rengong zhineng fuwu guanli...
- HiDreamv2HiDream most commonly refers to HiDream-I1, an open-source text-to-image generative foundation model released in April 2025 by the Chinese company HiDream.ai...
- MetaXv2MetaX (Chinese: 沐曦; pinyin: Mùxī), formally MetaX Integrated Circuits (Shanghai) Co., Ltd. and sometimes rendered Muxi, is a Chinese fabless semiconductor...
- Aquila (language model)v2Aquila (Chinese: 悟道·天鹰, Wudao Tianying) is a series of open bilingual (Chinese and English) large language models developed by the Beijing Academy of...
- Huawei PanGuv2PanGu (盘古, sometimes written Pangu) is a family of large AI models developed by Huawei, primarily through Huawei Cloud and Huawei's Noah's Ark Lab (诺亚方舟实验室)....
- CogAgentv2CogAgent is an open visual language model built to act as a graphical user interface (GUI) agent: given a screenshot and a natural-language goal, it predicts...
- CogVLMv2CogVLM is an open vision language model developed by Zhipu AI and the Knowledge Engineering Group (KEG) at Tsinghua University. It was introduced in the paper...
- Marco-o1v2Marco-o1 is an open reasoning model released in November 2024 by the MarcoPolo team at Alibaba International Digital Commerce (AIDC). It was introduced in the...
- CodeGeeXv3CodeGeeX is an open series of multilingual code generation models developed by the Knowledge Engineering Group (KEG) and Data Mining lab at Tsinghua University...
- DeepSeek-VLv2DeepSeek-VL is the first open-source vision-language model series from DeepSeek, the Chinese AI company. It was released on 11 March 2024 in 1.3B and 7B sizes,...
- DeepSeek LLMv2DeepSeek LLM is the first foundational large language model series released by the Chinese AI company DeepSeek. It was published on 29 November 2023 in two...
- Hierav2Hiera is a hierarchical vision transformer from Meta AI (FAIR), introduced in the paper "Hiera: A Hierarchical Vision Transformer without the...
- MEGABYTEv2MEGABYTE is a transformer architecture for autoregressive modeling of very long sequences directly at the byte level, introduced by researchers at Meta AI...
- LLM Compiler (Meta)v2The Meta Large Language Model Compiler, usually shortened to LLM Compiler, is a family of pre-trained large language model models built by Meta AI for code and...
- MetaCLIPv2MetaCLIP (Metadata-Curated Language-Image Pre-training) is a data curation recipe and a family of vision-language models from Meta AI, introduced in the 2023...
- Droidletv2Droidlet is an open-source platform from Facebook AI Research (now Meta AI) for building embodied AI agents. Released in 2021, it is a modular framework that...
- Grand Teton (AI hardware)v3Grand Teton is an open GPU hardware platform designed by Meta for training and running large AI models. Meta announced it at the 2022 Open Compute Project...
- Catalina (Meta AI rack)v2Catalina is a high-power, liquid-cooled rack system designed by Meta for training and serving large AI models. Meta unveiled it at the Open Compute Project...
- CICERO (AI)v2CICERO is an AI agent built by Meta AI's Fundamental AI Research (FAIR) division that reached human-level performance in the board game Diplomacy. Meta...
- Open Catalyst Projectv2The Open Catalyst Project (OCP) is a research collaboration between Meta AI's Fundamental AI Research group (FAIR) and Carnegie Mellon University's Department...
- Meta Motivov2Meta Motivo is a behavioral foundation model for controlling a simulated humanoid body, released by Meta AI's Fundamental AI Research (FAIR) group on December...
- Ego-Exo4Dv2Ego-Exo4D is a large-scale, multimodal, multiview video dataset and benchmark suite for computer vision research on skilled human activity. Its defining...
- AI Habitatv2AI Habitat (usually just Habitat) is an open-source simulation platform for embodied AI research, developed primarily by Meta AI (the group then known as...
- Nougat (model)v2Nougat (Neural Optical Understanding for Academic Documents) is a document-understanding model from Meta AI that converts the rendered image of a document page...
- Detectron2v2Detectron2 is an open-source software library for object detection and image segmentation, built on PyTorch and developed by Facebook AI Research (FAIR), the...
- I-JEPAv2I-JEPA (Image-based Joint-Embedding Predictive Architecture) is a self-supervised learning method for computer vision developed by Meta AI. It was introduced...
- Audioboxv2Audiobox is a foundation research model for audio generation developed by Meta AI and its Fundamental AI Research (FAIR) group. Announced in late 2023 and...
- data2vecv2data2vec is a self-supervised learning framework from Meta AI (then Facebook AI Research) that applies the same training method to three different input types:...
- No Language Left Behind (NLLB)v2No Language Left Behind (NLLB) is a machine translation research project and model family from Meta AI, announced in July 2022. Its flagship system, NLLB-200,...
- ImageBindv2ImageBind is a multimodal model from Meta AI (its Fundamental AI Research lab) that learns a single joint embedding space across six different modalities:...
- CM3leonv2CM3leon (pronounced "chameleon") is a multimodal generative model from Meta AI, introduced in July 2023, that handles both text-to-image and image-to-text...
- Make-A-Scenev2Make-A-Scene is a text-to-image generation model published by Meta AI (then Meta AI Research) in 2022. Its central idea is that a user can guide an image not...
- Emu Editv2Emu Edit is an instruction-based image editing model from Meta AI, announced on November 16, 2023 alongside the text-to-video model Emu Video. It edits a...
- Large Concept Modelv2A Large Concept Model (LCM) is a research approach to language modeling, introduced by Meta AI's Fundamental AI Research (FAIR) group in December 2024, that...
- Byte Latent Transformerv2The Byte Latent Transformer (BLT) is a tokenizer-free large language model architecture introduced by researchers at Meta AI's Fundamental AI Research (FAIR)...
- Llama APIv2The Llama API is Meta's first-party hosted cloud service for running Llama models. Announced on April 29, 2025 at Meta's inaugural LlamaCon developer...
- Llama Stackv2Llama Stack is an open standardized framework created by Meta for building generative AI applications. It defines a set of standardized APIs (building blocks)...
- BlenderBotv2BlenderBot is a line of open-domain conversational agents built by Facebook AI Research (FAIR), the lab now known as Meta AI. Three main versions were released...
- Atlas (language model)v2Atlas is a retrieval-augmented language model developed by researchers at Meta AI (the group then known as Facebook AI Research, or FAIR). It was introduced in...
- Doubao Seedreamv4Doubao-Seedream is the family of text-to-image generation foundation models developed by the ByteDance Seed team and shipped through ByteDance's Doubao product...
- MMStarv3MMStar (Multi-modal Star) is a vision-language model evaluation benchmark consisting of 1,500 multimodal samples that were filtered from six pre-existing...
- AlphaFold-Multimerv3AlphaFold-Multimer is a deep learning system for predicting the three-dimensional structures of protein complexes, released by Google DeepMind in October 2021...
- InternVideov3InternVideo is a family of general-purpose video foundation models developed by OpenGVLab at the Shanghai Artificial Intelligence Laboratory in collaboration...
- DeepSeek-Proverv4DeepSeek-Prover is a family of open-weight large language models developed by Chinese AI laboratory DeepSeek for formal theorem proving in the Lean 4 proof...
- AI in journalismv3AI in journalism refers to the use of artificial intelligence, and especially machine learning and generative AI, in the gathering, production, distribution...
- Chain of Density promptingv3Chain of Density (CoD) is a prompting technique for abstractive text summarization with large language models, introduced in the 2023 paper "From Sparse to...
- Crusoe Energyv3Crusoe Energy Systems, Inc. (commonly branded as Crusoe) is a privately held American energy and computing infrastructure company headquartered in Denver,...
- BentoMLv3BentoML is an open-source Python framework for packaging, serving, and deploying machine learning and AI models as production inference services. It was first...
- Oriol Vinyalsv4Oriol Vinyals (born 1983, Sabadell, Catalonia, Spain) is a Spanish machine learning researcher who serves as Vice President of Research at Google DeepMind and...
- 4NE-1v5This article is a detailed summary. For an even fuller treatment, see also NEURA Robotics 4NE-1. --- NEURA Robotics Type Germany (Metzingen,...
- Pallas (JAX kernel language)v3Pallas is an experimental extension to JAX that lets users write custom hardware kernels in Python and lower them to both Tensor Processing Units and NVIDIA...
- Candle (HuggingFace Rust ML)v3Candle is a minimalist machine learning framework written in pure Rust and published by Hugging Face under the huggingface/candle GitHub repository.[^1] The...
- Mike Lewisv3Mike Lewis is a British natural language processing researcher based in Seattle who serves as a research scientist at Meta AI (Facebook AI Research, FAIR) and...
- EAGLE-2v3EAGLE-2 ("Faster Inference of Language Models with Dynamic Draft Trees") is the second generation of the EAGLE family of speculative decoding methods for...
- Lookahead Decodingv3Lookahead Decoding is a parallel decoding algorithm for accelerating inference in large language models, introduced in November 2023 by Yichao Fu, Peter...
- OpenFoldv3OpenFold is an open source, trainable, GPU friendly PyTorch reimplementation of AlphaFold 2, developed initially by the AlQuraishi lab at Columbia University...
- Intel Loihiv3Intel Loihi is a family of research neuromorphic processors developed by Intel Labs to implement spiking neural networks (SNNs) in silicon, with the stated...
- Lightmatterv3Lightmatter is a U.S. silicon-photonics company that designs and manufactures optical computing hardware and photonic interconnects for artificial intelligence...
- DatologyAIv3DatologyAI is a Redwood City, California artificial-intelligence startup that builds automated tools for curating, deduplicating, and composing the training...
- Homomorphic encryption for machine learningv3Homomorphic encryption for machine learning is the application of fully, somewhat, or leveled homomorphic encryption (FHE, SHE, LHE) so that a server can run...
- Backdoor attacks on large language modelsv3A backdoor attack on a large language model (LLM) is an adversarial training-time attack in which an attacker manipulates training data, fine-tuning data,...
- Imbuev4Imbue is a San Francisco artificial intelligence research lab focused on training foundation models and building agent systems oriented toward reasoning and...
- Membership Inference Attackv3A Membership Inference Attack (MIA) is a privacy attack against a trained machine learning model in which an adversary, given a candidate data record and...
- OpenOrcav3OpenOrca is a large open-source instruction-tuning dataset that augments the FLAN Collection with chain-of-thought responses generated by OpenAI's GPT-3.5 and...
- Lumierev3Lumiere is a text-to-video diffusion model developed by Google Research in collaboration with researchers from the Weizmann Institute of Science, Tel Aviv...
- LTX-Videov3LTX-Video is an open-source, transformer-based latent video diffusion model developed by the Israeli company Lightricks and first released to the public in...
- H2O (Heavy-Hitter Oracle for KV Cache)v3H2O (Heavy-Hitter Oracle) is a training-free, runtime KV cache eviction policy for autoregressive large language model inference. It identifies a small subset...
- Instructor (library)v3Instructor is an open-source Python library that returns type-safe, Pydantic-validated structured outputs from large language model APIs. It works by wrapping...
- Guidance (library)v3Guidance is an open-source Python library, originally developed at Microsoft Research, for building structured, multi-step programs that drive large language...
- Outlines (library)v3Outlines is an open-source Python (programming language) library, released under the Apache 2.0 license, that constrains large language model output to...
- Open Interpreterv3Open Interpreter is an open-source desktop agent, distributed as a Python command-line tool and library, that lets a large language model write and execute...
- Kyutaiv3Kyutai is a privately funded nonprofit artificial intelligence research laboratory based in Paris, France, founded in November 2023 with an initial commitment...
- Latent Consistency Models (LCM)v4Latent Consistency Models (LCMs) are a family of accelerated text-to-image generative models that apply the consistency-models framework of Song et al. (2023)...
- NVIDIA L4v2The NVIDIA L4 is a compact, power-efficient data-center GPU built on the Ada Lovelace architecture and optimized for artificial-intelligence inference and...
- Humanoid robot autonomy levelsv4Humanoid robot autonomy levels describe a graduated framework for classifying the degree of independent operation that a humanoid robot can achieve. Ranging...
- ChipNeMov2ChipNeMo is a research project and a family of domain-adapted large language models developed by Nvidia to assist with industrial semiconductor and chip-design...
- NVIDIA Rubin Ultrav2NVIDIA Rubin Ultra is a planned data-center GPU platform from Nvidia, positioned on the company's roadmap as the mid-cycle "Ultra" refresh of the Rubin...
- LongLoRAv3LongLoRA is a parameter-efficient fine-tuning technique that extends the context window of pre-trained large language models with substantially lower...
- Limitless AIv3Limitless AI is an American personal artificial intelligence company best known for the Limitless Pendant, a small clip-on wearable that captures ambient...
- Boomyv4Boomy is a generative artificial intelligence music platform that lets people without musical training assemble original songs in a web browser and publish...
- NVIDIA Feynmanv2NVIDIA Feynman is the data center GPU architecture that NVIDIA has placed on its public roadmap as the successor to Rubin and Rubin Ultra, with a target of...
- AG-UI Protocolv3The AG-UI Protocol (Agent-User Interaction Protocol) is an open, event-based protocol that standardizes how AI agents communicate with user-facing...
- GLM-4-Voicev3GLM-4-Voice is an open-weights end-to-end speech-to-speech large language model released in October 2024 by Zhipu AI together with the Knowledge Engineering...
- OctoAIv3OctoAI (originally OctoML) was an American artificial intelligence infrastructure company that operated a generative-AI inference platform and, before its...
- fal.aiv3fal.ai (legally Features and Labels, Inc., often stylized fal) is a San Francisco-based artificial intelligence company that operates a specialized inference...
- Palisade Researchv3Palisade Research is a United States 501(c)(3) nonprofit research organization that studies the offensive capabilities of contemporary artificial intelligence...
- Normal Computingv4Normal Computing is an American deep-tech startup developing thermodynamic computing hardware and AI-driven electronic design automation (EDA) software....
- AG2 (framework)v2AG2 (formerly AutoGen) is an open-source Python framework for building applications composed of multiple cooperating large-language-model (LLM) agents. The...
- Magentic-Onev2Magentic-One is a generalist multi-agent system released by Microsoft Research's AI Frontiers lab in November 2024 to autonomously solve complex, multi-step...
- AI safety via debatev2AI safety via debate is a proposed approach to scalable oversight in which two artificial agents take turns presenting short statements about a question or...
- Model organisms of misalignmentv4Model organisms of misalignment is a research methodology in Anthropic's alignment program, and more broadly in AI-safety science, that calls for building...
- Eliciting latent knowledgev3Eliciting latent knowledge (ELK) is an open problem in AI alignment formulated by Paul Christiano, Ajeya Cotra, and Mark Xu at the Alignment Research Center...
- AI controlv3AI control is a research paradigm in technical AI safety that designs and evaluates deployment-time safety protocols under the explicit assumption that the...
- Instrumental convergencev4Instrumental convergence is a hypothesis in AI safety holding that a wide range of sufficiently capable agents, when pursuing almost any final goal, will...
- Flash-Decodingv2Flash-Decoding is an inference-time variant of the FlashAttention algorithm that targets the decoding (autoregressive generation) phase of large language model...
- CasiVisionv7--- Also known as Robotics, Humanoid robots Founded Haidian District, Beijing, China Key people CASIVIBOT, CASBOT 01, CASBOT 02, CASBOT W1, CASBOT...
- LawZerov3LawZero is a nonprofit AI safety research organization founded by Turing Award-winning computer scientist Yoshua Bengio and publicly launched on June 3, 2025....
- NVIDIA DSXv2NVIDIA DSX is a platform from NVIDIA for designing, simulating, building and operating large-scale data centers that the company calls AI factories. NVIDIA...
- Kadrey v. Metav2Kadrey v. Meta Platforms, Inc. is a putative class action lawsuit filed in 2023 by a group of book authors against Meta Platforms, Inc., alleging that Meta...
- nnsightv2nnsight is an open-source Python library for the interpretation and intervention of deep learning models, developed by the Bau Lab at Northeastern University....