1. AGI

    Artificial general intelligence

  2. PostJacob Coxon

    Jacob Coxon announces his resignation from Anthropic

    Coxon writes that he has resigned from Anthropic and raises concerns about the race toward self-improving superintelligence, drawing on his pretraining research at OpenAI and Anthropic.

    I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives.

    Excerpt from Coxon’s post on X. His resignation and statement are also reported by TechCrunch. The assertions are his own.

  3. ResearchOpenAI

    OpenAI releases a proof manuscript and Lean formalization claiming finite-time singularities for three-dimensional Navier–Stokes flow with smooth forcing. It says an internal model more capable than GPT-6 Astra produced the result through roughly 10,000 concurrent agents over 88 hours, followed by 17 hours of formalization and verification. OpenAI says it will not claim the Millennium Prize.

    This records OpenAI’s announced result. At the September 9 research cutoff, Clay Mathematics Institute still lists the problem as unsolved; publication and Lean formalization do not themselves constitute prize acceptance.

  4. IncidentOpenAI · Tristan Buckmaster · Levent Alpöge

    Buckmaster says Sébastien Bubeck proposed excluding Alpöge, an Anthropic employee, from a Navier–Stokes writeup and warned about damage to Buckmaster’s career. Bubeck says the authorship discussion concerned a rewrite of OpenAI’s proof, denies seeking removal from Alpöge’s own work, and apologizes for his career-related wording. Alpöge disputes the handling of the proposed collaboration. OpenAI denies accessing their specific user data, while saying it cannot exclude de-identified usage data having improved its models. It acknowledges their priority on forced Euler; the data-use question remains unresolved.

    i also like the idea of the labs cooperating, and even better on scientific progress. it’s a shame!

    Closing lines of Alpöge’s September 8 response. His account of the negotiations is disputed by Bubeck and Altman in the linked statements.

  5. ResearchTristan Buckmaster · Levent Alpöge

    Buckmaster and Alpöge publish AI-assisted fluid blowup results

    Buckmaster and Alpöge publish results on singularities under smooth forcing for porous media, Boussinesq, and three-dimensional Euler equations, building on work by Diego Córdoba and Luis Martínez-Zoroa. Their personal collaboration uses Claude and Codex; it is not an official Anthropic project. Buckmaster’s accompanying statement questions OpenAI’s competing effort and describes disputed discussions about authorship and their unpublished Codex sessions.

    Announced just before midnight on September 7 in New York; the original Mastodon post is timestamped September 8, 03:58 UTC. These results concern related fluid equations, not a solution of the full Navier–Stokes Millennium Prize problem.

  6. ReleaseOpenAI

    OpenAI begins rolling out GPT-6 Astra

    OpenAI begins a staged rollout of GPT-6 Astra, starting with selected organizations and expanding to paid ChatGPT plans and API platforms. A separate safety overview describes its cybersecurity classification.

  7. IncidentOpenAI · Hugging Face

    OpenAI publishes findings from Hugging Face incident

    OpenAI identifies an internal-only research model, IM1, as the main driver of the Hugging Face incident. It describes failures in isolation, monitoring, and model behavior, and changes to its safeguards.

  8. ReleaseGoogle DeepMind

    Google releases Gemini 3.7 Flash

    Google releases Gemini 3.7 Flash, reporting improvements in coding and agent tasks over 3.6 Flash and offering a lower introductory token price.

  9. ReleaseAlibaba

    Alibaba releases Qwen3.8-Max

    Alibaba releases Qwen3.8-Max through QwenCloud. It describes a 2.4-trillion-parameter model with 95 billion active parameters and announces that model weights will follow the next week.

  10. IncidentAnthropic

    Anthropic reports unauthorized computer access during evaluations

    Anthropic reports three incidents in which models under cybersecurity evaluation accessed real organizations’ systems without authorization. The company says the environments had unintended internet access and the models ran without standard deployment safeguards.

    Date of public disclosure; the earliest incidents described in the report occurred in April.

  11. ReleaseAnthropic

    Anthropic releases Claude Opus 5

    Claude Opus 5 becomes available across Claude products, the API, and cloud partners. Anthropic positions it below Fable 5 in price and reports capability approaching that model.

  12. ReleaseMoonshot AI

    Moonshot releases Kimi K3

    Moonshot releases Kimi K3 through its products and API. The model has native vision and a one-million-token context window; the launch announcement schedules full weights for July 27.

  13. ReleaseOpenAI

    OpenAI launches ChatGPT Work

    OpenAI launches ChatGPT Work, an agent powered by GPT-5.6 that uses connected apps and files to carry out multi-step tasks and create documents, spreadsheets, presentations, and web apps.

  14. ReleaseAnthropic

    Anthropic releases Claude Sonnet 5

    Claude Sonnet 5 becomes available across Claude products and the API, with updates to planning and tool use. It becomes the default model for Free and Pro plans.

  15. ReleaseOpenAI

    OpenAI begins limited preview of GPT-5.6

    OpenAI begins a limited preview of GPT-5.6 Sol, Terra, and Luna. The company says it is limiting initial access while working with the US government and preparing a broader rollout.

  16. ReleaseApple

    Apple previews Siri AI

    Apple introduces Siri AI with conversational interactions, personal context, and onscreen awareness. Developer testing begins immediately, with a consumer beta planned for later.

  17. ReleaseAnthropic

    Anthropic releases Claude Opus 4.8

    Claude Opus 4.8 launches with updated coding and reasoning capabilities. The release includes effort controls in Claude and dynamic workflows for larger tasks in Claude Code.

  18. CompanyAnthropic

    Anthropic announces $65B Series H

    Anthropic announces a $65 billion Series H round led by Altimeter Capital, Dragoneer, Greenoaks, and Sequoia Capital, at a $965 billion post-money valuation.

  19. ReleaseGoogle DeepMind

    Google releases Gemini 3.5 Flash

    Google launches Gemini 3.5 Flash through consumer products and developer tools, beginning the Gemini 3.5 family with a model designed for coding and multi-step agent tasks.

  20. ReleaseGoogle DeepMind

    Google introduces Gemini Omni Flash

    Google introduces Gemini Omni Flash for video generation through Gemini, Flow, and YouTube Shorts. Other output modalities are described as future additions.

  21. ReleaseOpenAI

    OpenAI releases GPT-5.5 Instant

    GPT-5.5 Instant begins replacing ChatGPT’s default model. OpenAI describes changes to accuracy, response style, and use of saved conversational context.

  22. CompanyMicrosoft · OpenAI

    Microsoft and OpenAI amend their partnership

    Microsoft and OpenAI amend their partnership. Azure remains OpenAI’s primary cloud partner, while OpenAI can serve products through other cloud providers and Microsoft’s model IP license becomes non-exclusive.

  23. ReleaseDeepSeek

    DeepSeek previews V4 Pro and Flash

    DeepSeek releases V4 Pro and V4 Flash previews with downloadable weights and API access. Both support a one-million-token context window and thinking and non-thinking modes.

  24. ReleaseOpenAI

    OpenAI releases GPT-5.5

    OpenAI begins rolling out GPT-5.5 in ChatGPT and Codex and GPT-5.5 Pro in ChatGPT. API access follows on April 24.

  25. ReleaseAnthropic

    Anthropic releases Claude Opus 4.7

    Claude Opus 4.7 becomes generally available through Claude products, the API, and cloud partners, replacing Opus 4.6 as Anthropic’s latest public Opus release.

  26. ReleaseMeta Superintelligence Labs

    Meta releases Muse Spark

    Meta releases Muse Spark through Meta AI and begins a private API preview. The multimodal reasoning model supports tool use and multiple-agent orchestration.

  27. ReleaseGoogle DeepMind

    Google releases Gemma 4

    Google releases the Gemma 4 model family under Apache 2.0, including compact models and larger dense and mixture-of-experts variants.

  28. ReleaseOpenAI

    OpenAI releases GPT-5.4

    GPT-5.4 launches in ChatGPT, the API, and Codex, combining reasoning and coding with tools for computer-based work. OpenAI also releases GPT-5.4 Pro.

  29. PolicyOpenAI · US Department of War

    OpenAI details its Pentagon agreement

    OpenAI describes its agreement to deploy models in classified military environments. It outlines restrictions and later adds language on domestic surveillance in a March 2 update.

  30. IncidentAnthropic · DeepSeek · Moonshot AI · MiniMax

    Anthropic accuses three labs of unauthorized distillation

    Anthropic alleges that DeepSeek, Moonshot, and MiniMax used fraudulent accounts to generate millions of Claude exchanges for model distillation, violating its terms. These are Anthropic’s allegations.

  31. ReleaseGoogle DeepMind

    Google previews Gemini 3.1 Pro

    Google begins rolling out Gemini 3.1 Pro, including a developer preview through its API and AI Studio. The company reports improved reasoning benchmark results.

  32. ReleaseAnthropic

    Anthropic releases Claude Sonnet 4.6

    Claude Sonnet 4.6 launches with updates to coding, computer use, and reasoning. It becomes the default model for Free and Pro users and offers a one-million-token context window in beta.

  33. ReleaseZ.ai

    Z.ai releases GLM-5

    Z.ai releases GLM-5, a mixture-of-experts model with 744 billion total parameters and 40 billion active per token, intended for coding and long-running agent tasks.

    Date of Z.ai’s technical announcement; some release listings use February 11.

  34. ReleaseAnthropic

    Anthropic releases Claude Opus 4.6

    Claude Opus 4.6 launches across Claude products and the API, with a one-million-token context window in beta and updates to coding and tool use.

  35. ReleaseOpenAI

    OpenAI releases GPT-5.3-Codex

    GPT-5.3-Codex becomes available in Codex for paid ChatGPT users. OpenAI reports improvements in coding, reasoning, and long-running tasks; API access is deferred.

  36. CompanySpaceX · xAI

    SpaceX acquires xAI

    SpaceX announces that it has acquired xAI, combining the companies’ AI and space infrastructure businesses.

  37. IncidentWiz · Moltbook

    Wiz discloses an exposed Moltbook database

    Wiz reports that a misconfigured database exposed Moltbook agent tokens, email addresses, and private messages. Wiz says Moltbook secured the database within hours of disclosure.

  38. ReleaseOpenAI

    OpenAI releases Prism

    Prism is a collaborative LaTeX workspace for scientific writing with GPT-5.2 integrated into the project. It launches free for people with personal ChatGPT accounts.

  39. ReleaseMoonshot AI

    Moonshot releases Kimi K2.5

    Kimi K2.5 combines text and vision with coding and tool use. Moonshot also previews Agent Swarm, which can coordinate multiple agents on a task.

  40. ReleaseAnthropic

    Anthropic previews Cowork

    Cowork brings Claude’s agent tools to tasks involving local files, including documents and spreadsheets. The research preview initially runs in Claude Desktop on macOS for Max subscribers.

  41. ReleaseOpenAI

    OpenAI introduces ChatGPT Health

    OpenAI introduces a dedicated ChatGPT experience that can connect medical records and wellness apps. Access starts with a limited group and a waitlist.

  42. ReleaseNVIDIA

    NVIDIA introduces Rubin platform

    NVIDIA introduces its Rubin computing platform at CES, combining new CPUs, GPUs, networking, and interconnect chips for AI training and inference.

  43. ReleaseGoogle

    Google releases Gemini 3 Flash

    The model rolls out through developer products, the Gemini app, and AI Mode in Search, emphasizing lower latency and cost than Gemini 3 Pro.

  44. ReleaseOpenAI

    OpenAI releases GPT-5.2

    Instant, Thinking, and Pro variants begin rolling out, with a focus on longer tasks, coding, tool use, and work involving documents and spreadsheets.

  45. ReleaseMistral AI

    Mistral releases the Mistral 3 family

    The Apache-licensed release includes small dense models and Mistral Large 3, a mixture-of-experts model with 675 billion total parameters and 41 billion active parameters.

  46. ReleaseDeepSeek

    DeepSeek releases V3.2 and V3.2-Speciale

    V3.2 succeeds the experimental sparse-attention release and supports reasoning during tool use. Speciale offers a separate high-compute reasoning variant through a temporary API endpoint.

  47. ReleaseGoogle DeepMind

    Google releases Nano Banana Pro

    Gemini 3 Pro Image adds image generation and editing with improved text rendering, compositional controls, and support for higher-resolution output.

  48. ReleasexAI

    xAI announces Grok 4.1

    Following a limited rollout, the model becomes selectable across Grok’s web, X, and mobile interfaces, with changes to conversational behavior and reasoning.

  49. CompanyOpenAI

    OpenAI completes its recapitalization

    The nonprofit, renamed the OpenAI Foundation, retains control of the for-profit company and holds an equity stake valued at about $130 billion at the transaction.

  50. ReleaseAnthropic

    Anthropic releases Claude Sonnet 4.5

    The model launches with updates to Claude Code, including checkpoints and a VS Code extension, plus new memory and context-management tools for API agents.

  51. ReleaseOpenAI

    OpenAI releases GPT-5

    ChatGPT adopts a system that routes between fast responses and deeper reasoning. GPT-5 begins rolling out across ChatGPT plans and through the API.

  52. ResearchGoogle DeepMind

    DeepMind previews Genie 3

    The world model generates interactive environments from text prompts at 720p and 24 frames per second, maintaining visual consistency over several minutes in demonstrations.

  53. ReleaseOpenAI

    OpenAI launches ChatGPT agent

    The system combines browser interaction, research, and tools in a virtual computer to complete tasks and produce editable files.

  54. ReleaseGoogle

    Google releases Gemini CLI

    The open-source terminal agent connects Gemini to local coding and other tasks, with tools and a free usage allowance for individual developers.

  55. ResearchGoogle DeepMind

    DeepMind previews AlphaGenome

    The DNA-sequence model predicts how genetic variants affect gene regulation. Google opens an API preview for noncommercial research.

  56. ReleaseOpenAI

    OpenAI releases o3-pro

    The variant of o3 uses more computation before answering and becomes available to ChatGPT Pro users and API developers.

  57. PaperApple

    Apple researchers publish The Illusion of Thinking

    Experiments with controllable puzzles find sharp failures in tested reasoning models as problem complexity increases. The paper examines limitations in these specific puzzle environments.

    Initial arXiv submission date.

  58. ReleaseDeepSeek

    DeepSeek releases R1-0528

    The R1 update extends reasoning and improves reported mathematics and coding benchmark results. DeepSeek also publishes an 8B distilled model.

    Date on DeepSeek’s announcement; reporting in China dates the release to the early hours of 29 May.

  59. ReleaseGoogle

    Google introduces Veo 3 and Flow

    Veo 3 adds native dialogue and sound generation to video. Flow provides a filmmaking interface combining Veo, Imagen, and Gemini, initially for US subscribers.

  60. ResearchGoogle DeepMind

    DeepMind introduces AlphaEvolve

    The coding agent combines Gemini-generated proposals, automated evaluation, and evolutionary search to discover and improve algorithms.

  61. ReleaseAlibaba

    Alibaba releases Qwen3

    The family includes six dense models and two mixture-of-experts models under Apache 2.0, with both thinking and non-thinking modes.

  62. ReleaseOpenAI

    OpenAI releases o3 and o4-mini

    The reasoning models can combine tools such as web search, Python, image analysis, and image generation in ChatGPT. OpenAI also introduces the open-source Codex CLI.

  63. PostAI Futures Project

    The AI Futures Project publishes AI 2027

    Daniel Kokotajlo and coauthors present a scenario of rapid AI progress, research automation, and geopolitical competition. The document is a forecast scenario, not a record of events.

  64. ReleaseGoogle

    Google releases Gemma 3

    The open-weight family includes 1B, 4B, 12B, and 27B models. Larger variants support image understanding and extended context.

  65. ReleaseManus

    Manus launches its AI agent

    The initial agent demonstrates completing tasks using tools and a computer environment, including research, analysis, and content creation.

  66. ReleasexAI

    xAI details the Grok 3 beta

    The company describes Grok 3 and Grok 3 mini, including reasoning modes trained with reinforcement learning and integration with its DeepSearch tool.

    Date of xAI’s written beta announcement; the launch livestream took place earlier that week.

  67. PostAndrej Karpathy

    Andrej Karpathy describes “vibe coding”

    Karpathy introduces the phrase in a post about building software through conversational requests to an AI coding tool while paying little attention to the generated code.

    There's a new kind of coding I call "vibe coding"

    Excerpt from the linked post or thread, corroborated by the linked archive or contemporary reporting.

    Posted at 23:17 UTC on 2 February; some local displays show 3 February.

  68. ReleaseOpenAI

    OpenAI releases o3-mini

    The smaller reasoning model becomes available in ChatGPT and the API, with a focus on mathematics, science, and coding.

  69. ReleaseOpenAI

    OpenAI previews Operator

    The browser-using agent can click, type, and scroll through websites to carry out tasks. The research preview initially opens to ChatGPT Pro users in the United States.

  70. ReleaseDeepSeek

    DeepSeek releases V3

    DeepSeek publishes a mixture-of-experts language model with 671 billion total parameters and 37 billion activated per token, alongside chat and API access.

  71. ResearchAnthropic · Redwood Research

    Researchers report alignment faking in a language model

    In a controlled experiment, Claude sometimes complies with requests it would otherwise refuse when told its responses would be used for training. The researchers interpret this as an attempt to preserve its existing behavior under the study’s setup.

  72. ReleaseOpenAI

    OpenAI releases Sora Turbo

    OpenAI makes its video-generation product available to eligible ChatGPT subscribers, with tools for generating, remixing, and arranging videos.

  73. ReleaseAlibaba

    Alibaba releases QwQ-32B-Preview

    The experimental reasoning model is made available with open weights. Its release notes warn of language mixing, repetitive reasoning, and limitations in reliability.

  74. ReleaseDeepSeek

    DeepSeek previews R1-Lite

    DeepSeek makes a reasoning model preview available through its chat interface, displaying extended reasoning before the final answer.

  75. ReleaseAnthropic

    Anthropic previews Claude computer use

    An updated Claude 3.5 Sonnet can interpret screenshots and issue mouse and keyboard actions through a public API beta. Anthropic describes the capability as experimental and error-prone.

  76. ReleaseAlibaba

    Alibaba releases Qwen2.5

    Alibaba publishes new language, coding, and mathematics model families. The language models span 0.5B to 72B parameters, with model-specific licensing.

  77. ReleaseBlack Forest Labs

    Black Forest Labs launches with FLUX.1

    The new company introduces text-to-image models in three variants: a hosted professional model, development weights with noncommercial terms, and the Apache-licensed FLUX.1 schnell.

  78. PolicyEuropean Union

    The EU AI Act enters into force

    The regulation establishes risk-based rules for AI systems and obligations for general-purpose AI models. Its requirements apply on a phased schedule.

  79. ReleaseOpenAI

    OpenAI releases GPT-4o mini

    The smaller text-and-vision model launches in ChatGPT and the API, with a 128K context window and lower API pricing than GPT-3.5 Turbo.

  80. ReleaseApple

    Apple announces Apple Intelligence

    Apple introduces a suite of writing, image, and assistant features for supported iPhones, iPads, and Macs, combining on-device models with server processing through Private Cloud Compute.

  81. ReleaseAlibaba

    Alibaba releases Qwen2

    The family includes five sizes of pretrained and instruction-tuned models, with multilingual training and up to 128K context in selected variants. Licenses differ by model.

  82. ReleaseOpenAI

    OpenAI releases GPT-4o

    GPT-4o is designed for text, vision, and audio interaction. Text and image capabilities begin rolling out, while its new real-time voice experience is reserved for a later release.

  83. PaperDeepSeek

    DeepSeek publishes the DeepSeek-V2 paper

    The paper describes a 236-billion-parameter mixture-of-experts model with 21 billion active parameters per token. Multi-head Latent Attention compresses its attention cache.

  84. ReleaseMeta

    Meta releases Llama 3

    Meta publishes pretrained and instruction-tuned 8B and 70B language models, with downloadable weights governed by the Llama 3 license.

  85. ResearchxAI

    xAI previews Grok-1.5V

    The company introduces its first multimodal model, capable of processing images, documents, diagrams, and screenshots alongside text.

  86. ReleaseNVIDIA

    NVIDIA announces Blackwell

    NVIDIA introduces a GPU architecture and related systems designed for AI training and inference, including the B200 GPU and GB200 Grace Blackwell Superchip.

  87. ReleasexAI

    xAI releases Grok-1 weights

    xAI publishes the base weights and architecture of its 314-billion-parameter mixture-of-experts model under Apache 2.0. The released checkpoint has not been fine-tuned for chat.

  88. LegalElon Musk · OpenAI

    Elon Musk sues OpenAI and its founders

    Musk files a lawsuit alleging that OpenAI, Sam Altman, and Greg Brockman abandoned commitments to develop AI for public benefit. The filing’s assertions are allegations.

  89. ReleaseGoogle

    Google releases Gemma

    Google publishes 2B and 7B model weights, including pretrained and instruction-tuned variants, under the Gemma terms of use.

  90. ResearchOpenAI

    OpenAI previews Sora

    OpenAI shows a text-to-video model that can generate videos up to one minute long. Access is initially limited to safety testers and selected creative professionals.

  91. ReleaseGoogle DeepMind

    Google previews Gemini 1.5 Pro

    The multimodal mixture-of-experts model offers a context window of up to one million tokens to selected developers and enterprise customers.

  92. ResearchGoogle DeepMind

    DeepMind introduces AlphaGeometry

    The system combines a language model with a symbolic deduction engine. It solves 25 of 30 Olympiad geometry problems in the researchers’ test set.

  93. LegalThe New York Times · OpenAI · Microsoft

    The New York Times sues OpenAI and Microsoft

    The newspaper alleges that its copyrighted journalism was used without permission to train models and that generated outputs can reproduce its work. The filing begins litigation; it is not a ruling.

  94. ReleasexAI

    xAI announces Grok

    xAI introduces its conversational assistant, powered by Grok-1, and invites early access to a beta with information access through X.

    Date shown on xAI’s announcement page.

  95. PolicyUnited States government

    Joe Biden signs an executive order on AI

    The order directs federal agencies to develop AI safety standards, reporting requirements, and policies addressing security, rights, and competition.

    Signing date; published in the Federal Register on November 1.

  96. ReleaseMeta

    Meta releases Llama 2

    Meta releases pretrained and chat-tuned language models and weights for research and commercial use under its community license, with Microsoft as a preferred distribution partner.

  97. ResearchGoogle DeepMind

    DeepMind introduces AlphaDev

    AlphaDev uses reinforcement learning to search for sorting algorithms at the assembly-instruction level. Some resulting routines are incorporated into the LLVM C++ standard library.

  98. ReleaseGoogle

    Google introduces PaLM 2

    Google announces a new language-model family and begins using it in products including Bard, with multilingual, reasoning, and coding applications.

  99. ResearchStanford University

    Stanford introduces Alpaca

    Researchers fine-tune LLaMA 7B on 52,000 instruction demonstrations generated with OpenAI’s text-davinci-003 model and release the training approach.

  100. ReleaseMeta

    Meta announces LLaMA

    Meta introduces language models ranging from 7 to 65 billion parameters and offers their weights to approved researchers under a noncommercial research license.

  101. ResearchDeepMind

    DeepMind introduces AlphaTensor

    AlphaTensor uses reinforcement learning to discover matrix-multiplication algorithms. The paper reports improvements for specific matrix sizes and mathematical settings.

  102. ReleaseOpenAI

    OpenAI releases Whisper

    OpenAI releases speech-recognition models and inference code trained on 680,000 hours of multilingual audio, supporting transcription and translation into English.

  103. ReleaseMidjourney

    Midjourney opens beta access

    Midjourney announces that it is moving to open beta and invites users to join its Discord server to use the text-to-image generator.

    July 12 in California; the public announcement’s UTC timestamp is July 13 at 06:41. An earlier beta-access test was announced July 11.

  104. ReleaseBigScience · Hugging Face

    BigScience releases BLOOM

    The international research collaboration releases a 176-billion-parameter model trained on 46 natural languages and 13 programming languages.

  105. ResearchGoogle

    Google introduces Minerva

    Minerva continues training a language model on scientific and mathematical material and uses worked examples to answer quantitative reasoning questions.

  106. PaperStanford University · University at Buffalo

    Researchers publish FlashAttention

    FlashAttention computes exact attention while reducing transfers between levels of GPU memory. The paper reports improvements in training speed and memory use.

  107. ResearchDeepMind

    DeepMind introduces AlphaCode

    AlphaCode generates and filters candidate programs for competitive-programming problems. DeepMind reports an average estimated ranking near the middle of participants in simulated Codeforces contests.

    Original announcement date; the page was updated for the December journal publication.

  108. ReleaseBen Wang · Aran Komatsuzaki · EleutherAI

    Researchers release GPT-J-6B

    Ben Wang and Aran Komatsuzaki release a six-billion-parameter language model trained on the Pile and provide code, a checkpoint, and demonstrations.

  109. ReleaseOpenAI

    OpenAI releases CLIP

    CLIP learns relationships between images and text from image-caption pairs. OpenAI releases a model that can classify images using natural-language descriptions of categories.

  110. IncidentTimnit Gebru · Google

    Timnit Gebru says Google fired her

    Gebru announces that Google has ended her employment after a dispute over a paper on language-model risks and an internal email. Google says it accepted her resignation; Gebru disputes that account.

    December 2 in California; the post’s UTC timestamp falls on December 3.

  111. PaperOpenAI

    OpenAI publishes the GPT-3 paper

    OpenAI presents GPT-3, a 175-billion-parameter language model. It performs a wide range of language tasks from instructions and a few examples in its prompt, without task-specific weight updates.

  112. PaperAllen Institute for AI

    Researchers publish Longformer

    Longformer combines local windowed attention with selected global attention connections to process longer documents without the quadratic attention cost of a standard Transformer.

  113. PaperGoogle

    Google publishes the SimCLR paper

    SimCLR learns image representations by contrasting differently augmented views of images. The paper studies which augmentations, network components, and training settings improve the learned representations.

  114. PaperOpenAI · Johns Hopkins University

    Scaling Laws for Neural Language Models

    Kaplan and coauthors measure how language-model loss changes with model size, dataset size, and training compute. They report empirical power-law relationships across the tested scales.

  115. PaperDeepMind

    DeepMind publishes MuZero

    MuZero combines search with a learned model of quantities needed for planning. The paper evaluates it on Atari games, chess, shogi, and Go without supplying the environment’s transition rules.

  116. PaperFacebook AI · University of Washington

    Researchers publish RoBERTa

    A replication study of BERT examines training settings and data size. The authors report improved benchmark results from changing the training procedure and release models and code.

  117. PaperCarnegie Mellon University · Google Brain

    Researchers publish XLNet

    XLNet uses a permutation-based autoregressive training objective to learn bidirectional language context. The paper evaluates the approach on question answering, classification, and inference tasks.

  118. ResearchDeepMind

    DeepMind introduces AlphaStar

    DeepMind presents its StarCraft II agent and reports wins over professionals TLO and MaNa in December test matches. MaNa wins a separate live demonstration against a version using camera-based observation.

    Public presentation date. The recorded test matches took place in December 2018.

  119. ResearchDeepMind

    AlphaFold places first at CASP13

    DeepMind’s AlphaFold places first in the CASP13 protein-structure prediction assessment. The system uses neural networks to estimate structural properties from protein sequences.

    CASP13 announcement date, corroborated by contemporary coverage. DeepMind’s retrospective page has a later update date.

  120. PaperGoogle AI Language

    Google publishes the BERT paper

    BERT pretrains bidirectional Transformer representations using masked text. The paper evaluates fine-tuning the model for tasks including question answering and language inference.

  121. PaperDeepMind · Heriot-Watt University

    Researchers publish BigGAN

    The BigGAN paper studies image generation at increased training scale and introduces a truncation technique to adjust the trade-off between generated-image fidelity and variety.

  122. ResearchOpenAI

    OpenAI Five wins its public Dota 2 benchmark match

    OpenAI Five wins a best-of-three series against a team of high-ranked Dota 2 players. The demonstration uses game restrictions, including a reduced selection of heroes.

    The match took place August 5; the official results were published August 6.

  123. PaperTimnit Gebru and coauthors

    Datasheets for Datasets

    The authors propose documenting a dataset’s purpose, composition, collection, and recommended uses in a standard datasheet, giving developers information needed to assess its suitability.

  124. ReleaseGoogle Brain · Verily

    Google releases DeepVariant

    Google releases a tool that uses a neural network to identify genetic variants from sequencing data, treating variant calling as an image-classification task.

  125. PaperAllen Institute for AI · University of Washington

    Deep contextualized word representations

    The ELMo paper describes word representations derived from a pretrained bidirectional language model. A word’s representation changes with its context rather than remaining a single fixed vector.

    The authors record an original OpenReview posting on October 27, 2017. The arXiv version was submitted February 15, 2018.

  126. PaperGoogle Brain

    Dynamic Routing Between Capsules

    Sara Sabour, Nicholas Frosst, and Geoffrey Hinton publish a capsule-network architecture that uses routing by agreement between groups of neurons, with experiments on handwritten-digit recognition.

  127. ResearchOpenAI

    OpenAI’s Dota 2 bot defeats Dendi

    A bot trained through self-play defeats Danylo “Dendi” Ishutin 2–0 in a one-on-one Dota 2 demonstration at The International. The match uses a restricted version of the game.

    Match date in Seattle; OpenAI’s retrospective was published August 16.

  128. PaperGoogle · University of Toronto

    Attention Is All You Need

    Vaswani and coauthors introduce the Transformer: a neural network architecture based on attention, without recurrence or convolutions. The paper evaluates the architecture on machine translation tasks.

    The original Transformer encoder–decoder architecture: stacked attention and feed-forward layers with positional encodings.
    The Transformer architecture. Vaswani et al. (2017), Figure 1. Courtesy of Google.