AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
Model Releases

important (and very jakub-coded) jakub quote:

DGX agent

important (and very jakub-coded) jakub quote: OpenAI Unveils GPT-5.5. Company Says Expect a Faster Model Release Pace 👀 OpenAI: 'We see pretty significant improvements in the short term, extremely sig

model-releasessam-altman--x
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

🚨 In a new court filing (below), Clippers owner Steve Ballmer dismisses @pablofindsout as “gossip” from a “former talking head and televisi…

DGX agent

🚨 In a new court filing (below), Clippers owner Steve Ballmer dismisses @pablofindsout as “gossip” from a “former talking head and television personality.” Here is an excerpt from the federal whistleb

model-releasesanthropic--x
23 Apr 2026
Model Releases

In ChatGPT, full-stack inference improvements enable a more capable model at faster speed. This efficiency is a game-changer for GPT-5.5 Pro…

DGX agent

In ChatGPT, full-stack inference improvements enable a more capable model at faster speed. This efficiency is a game-changer for GPT-5.5 Pro, now a much more practical option for demanding tasks, and

model-releasesopenai--x
23 Apr 2026
Model Releases

Infection-Reasoner: A Compact Vision-Language Model for Wound Infection Classification with Evidence-Grounded Clinical Reasoning

DGX agent

arXiv:2604.19937v1 Announce Type: cross Abstract: Assessing chronic wound infection from photographs is challenging because visual appearance varies across wound etiologies, anatomical locations, and

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions (Sydney Bradley/Business Insider)

DGX agent

Sydney Bradley / Business Insider: Instagram launches Instants, an app for sharing disappearing photos, in Italy and Spain, after rolling out an Instants feature in its main app in some regions — - In

model-releasestechmeme
23 Apr 2026
Model Releases

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimi…

DGX agent

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimited time + web access on a benchmark of real & hard clinical

model-releasesethan-mollick--x
23 Apr 2026
Model Releases

Intersectional Fairness in Large Language Models

DGX agent

arXiv:2604.20677v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in socially sensitive settings, raising concerns about fairness and biases, particularly across i

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its w…

DGX agent

Introducing GPT-5.5 A new class of intelligence for real work and powering agents, built to understand complex goals, use tools, check its work, and carry more tasks through to completion. It marks a

model-releasesopenai--x
23 Apr 2026
Model Releases

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against c…

DGX agent

I've been previewing this in Codex for a few weeks - it's very good! Had some great results from it having it run security reviews against code written using other models Introducing GPT-5.5 A new cla

model-releasessimon-willison--x
23 Apr 2026
Model Releases

IVY-FAKE: A Unified Explainable Framework and Benchmark for Image and Video AIGC Detection

DGX agent

arXiv:2506.00979v5 Announce Type: replace-cross Abstract: The rapid development of Artificial Intelligence Generated Content (AIGC) techniques has enabled the creation of high-quality synthetic conten

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

KANMixer: a minimal KAN-centered mixer for long-term time series forecasting

DGX agent

arXiv:2508.01575v2 Announce Type: replace Abstract: Long-term time series forecasting (LTSF) underpins critical applications from energy management to weather prediction, yet achieving reliable multi-

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Kimi K2.6 becomes the #1 open model on MathArena!

DGX agent

Kimi K2.6 achieved the top ranking on MathArena, a benchmark for evaluating mathematical problem-solving capabilities in open-source language models. This announcement highlights the model's superior

model-releaseskimi-moonshot--x
23 Apr 2026
Model Releases

Knapsack Optimization-based Schema Linking for LLM-based Text-to-SQL Generation

DGX agent

arXiv:2502.12911v3 Announce Type: replace Abstract: Generating SQLs from user queries is a long-standing challenge, where the accuracy of initial schema linking significantly impacts subsequent SQL ge

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Knowledge Capsules: Structured Nonparametric Memory Units for LLMs

DGX agent

arXiv:2604.20487v1 Announce Type: cross Abstract: Large language models (LLMs) encode knowledge in parametric weights, making it costly to update or extend without retraining. Retrieval-augmented gene

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

KoALa-Bench: Evaluating Large Audio Language Models on Korean Speech Understanding and Faithfulness

DGX agent

arXiv:2604.19782v1 Announce Type: cross Abstract: Recent advances in large audio language models (LALMs) have enabled multilingual speech understanding. However, benchmarks for evaluating LALMs remain

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

DGX agent

arXiv:2601.13240v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain special

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Large Language Models Meet Biomedical Knowledge Graphs for Mechanistically Grounded Therapeutic Prioritization

DGX agent

arXiv:2604.19815v1 Announce Type: new Abstract: Drug repurposing is often framed as a candidate identification task, but existing approaches provide limited guidance for distinguishing biologically pl

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure

DGX agent

arXiv:2604.20652v1 Announce Type: new Abstract: Large language models trained on human feedback may suppress fraud warnings when investors arrive already persuaded of a fraudulent opportunity. We test

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Last night was the biggest disaster in the history of Tesla. Let me walk you through what actually happened on that earnings call, because t…

DGX agent

Last night was the biggest disaster in the history of Tesla. Let me walk you through what actually happened on that earnings call, because the headlines are doing you a disservice: Elon Musk got on th

model-releasesgary-marcus--x
23 Apr 2026
Model Releases

Last week, we launched Gemini 3.1 TTS, our latest and best text-to-speech model. This new model introduces [awe] audio tags, an intuitive wa…

DGX agent

Last week, we launched Gemini 3.1 TTS, our latest and best text-to-speech model. This new model introduces [awe] audio tags, an intuitive way to guide vocal style, pace, and delivery. Here are some ti

model-releasesgoogle-ai--x
23 Apr 2026
Model Releases

Latent Stochastic Interpolants

DGX agent

arXiv:2506.02276v2 Announce Type: replace Abstract: Stochastic Interpolants (SI) is a powerful framework for generative modeling, capable of flexibly transforming between two probability distributions

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

LayerTracer: A Joint Task-Particle and Vulnerable-Layer Analysis framework for Arbitrary Large Language Model Architectures

DGX agent

arXiv:2604.20556v1 Announce Type: cross Abstract: Currently, Large Language Models (LLMs) feature a diversified architectural landscape, including traditional Transformer, GateDeltaNet, and Mamba. How

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Learning to Evolve: A Self-Improving Framework for Multi-Agent Systems via Textual Parameter Graph Optimization

DGX agent

arXiv:2604.20714v1 Announce Type: new Abstract: Designing and optimizing multi-agent systems (MAS) is a complex, labor-intensive process of 'Agent Engineering.' Existing automatic optimization methods

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication

DGX agent

arXiv:2604.19895v1 Announce Type: new Abstract: A well-known limitation of AI systems is presumptuousness: the tendency of AI systems to provide confident answers when information may be lacking. This

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Less Languages, Less Tokens: An Efficient Unified Logic Cross-lingual Chain-of-Thought Reasoning Framework

DGX agent

arXiv:2604.20090v1 Announce Type: new Abstract: Cross-lingual chain-of-thought (XCoT) with self-consistency markedly enhances multilingual reasoning, yet existing methods remain costly due to extensiv

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent

DGX agent

arXiv:2604.17931v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has emerged as a powerful training paradigm for LLM-based agents. However, scaling agentic RL for deep research remains

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?

DGX agent

arXiv:2501.03624v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel on many NLP benchmarks, but their behavior on real-world, semi-structured prediction remains underexplored.

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

LLM-guided phase diagram construction through high-throughput experimentation

DGX agent

arXiv:2604.20304v1 Announce Type: cross Abstract: Constructing phase diagrams for multicomponent alloys requires extensive experimental measurements and is a time-consuming task. Here we investigate w

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m outp…

DGX agent

looks like new Pareto frontiers across everything: - Context: 400K context in Codex and a 1M in API - API Pricing: 5/m input and 30/m output tokens. - Codex improved its own inference speed 20% lol -

model-releasesswyx--x
23 Apr 2026
Model Releases

LoRA-FA: Efficient and Effective Low Rank Representation Fine-tuning

DGX agent

arXiv:2308.03303v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) is crucial for improving their performance on downstream tasks, but full-parameter fine-tuning (Full-FT) is

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

MambaLiteUNet: Cross-Gated Adaptive Feature Fusion for Robust Skin Lesion Segmentation

DGX agent

arXiv:2604.20286v1 Announce Type: cross Abstract: Recent segmentation models have demonstrated promising efficiency by aggressively reducing parameter counts and computational complexity. However, the

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MAPRPose: Mask-Aware Proposal and Amodal Refinement for Multi-Object 6D Pose Estimation

DGX agent

arXiv:2604.20650v1 Announce Type: new Abstract: 6D object pose estimation in cluttered scenes remains challenging due to severe occlusion and sensor noise. We propose MAPRPose, a two-stage framework t

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Measuring the Machine: Evaluating Generative AI as Pluralist Sociotechical Systems

DGX agent

arXiv:2604.20545v1 Announce Type: new Abstract: In measurement theory, instruments do not simply record reality; they help constitute what is observed. The same holds for generative AI evaluation: ben

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Meta-Tool: Efficient Few-Shot Tool Adaptation for Small Language Models

DGX agent

arXiv:2604.20148v1 Announce Type: cross Abstract: Can small language models achieve strong tool-use performance without complex adaptation mechanisms? This paper investigates this question through Met

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management

DGX agent

arXiv:2601.11505v2 Announce Type: replace-cross Abstract: Progress in Type 1 Diabetes (T1D) algorithm development is limited by the fragmentation and lack of standardization across existing T1D manage

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Microsoft launches ‘vibe working’ in Word, Excel, and PowerPoint

DGX agent

Microsoft is rolling out a new Agent Mode inside Office apps like Word, Excel, and PowerPoint this week. Previously described by Microsoft as 'vibe working,' the Agent Mode is a more powerful version

model-releasesthe-verge-ai
23 Apr 2026
Model Releases

MIRROR: A Hierarchical Benchmark for Metacognitive Calibration in Large Language Models

DGX agent

arXiv:2604.19809v1 Announce Type: new Abstract: We introduce MIRROR, a benchmark comprising eight experiments across four metacognitive levels that evaluates whether large language models can use self

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror

DGX agent

arXiv:2604.14785v2 Announce Type: replace Abstract: Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their poten

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation

DGX agent

arXiv:2604.20366v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) exhibit powerful generative capabilities but frequently produce hallucinations that compromise output reliability.

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Mitigating Prompt-Induced Cognitive Biases in General-Purpose AI for Software Engineering

DGX agent

arXiv:2604.16756v2 Announce Type: replace-cross Abstract: Prompt-induced cognitive biases are changes in a general-purpose AI (GPAI) system's decisions caused solely by biased wording in the input (e.

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MixLLM: LLM Quantization with Global Mixed-precision between Output-features and Highly-efficient System Design

DGX agent

arXiv:2412.14590v2 Announce Type: replace Abstract: Quantization has become one of the most effective methodologies to compress LLMs into smaller size. However, the existing quantization solutions sti

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

MLG-Stereo: ViT Based Stereo Matching with Multi-Stage Local-Global Enhancement

DGX agent

arXiv:2604.20393v1 Announce Type: new Abstract: With the development of deep learning, ViT-based stereo matching methods have made significant progress due to their remarkable robustness and zero-shot

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Model Capability Assessment and Safeguards for Biological Weaponization

DGX agent

arXiv:2604.19811v1 Announce Type: cross Abstract: AI leaders and safety reports increasingly warn that advances in model reasoning may enable biological misuse, including by low-expertise users, while

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

MSLAU-Net: A Hybrid CNN-Transformer Network for Medical Image Segmentation

DGX agent

arXiv:2505.18823v2 Announce Type: replace Abstract: Accurate medical image segmentation allows for the precise delineation of anatomical structures and pathological regions, which is essential for tre

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Mythos and the Unverified Cage: Z3-Based Pre-Deployment Verification for Frontier-Model Sandbox Infrastructure

DGX agent

arXiv:2604.20496v1 Announce Type: cross Abstract: The April 2026 Claude Mythos sandbox escape exposed a critical weakness in frontier AI containment: the infrastructure surrounding advanced models rem

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

New in the Codex app: - GPT-5.5 - Browser control - Sheets & Slides - Docs & PDFs - OS-wide dictation - Auto-review mode Enjoy!

DGX agent

The Codex app now includes several new features: GPT-5.5 integration, browser control capabilities, support for Google Sheets and Slides, document and PDF handling, OS-wide dictation functionality, an

model-releasessam-altman--x
23 Apr 2026
Model Releases

'Newspaper Eat' Means 'Not Tasty': A Taxonomy and Benchmark for Coded Language in Real-World Chinese Online Reviews

DGX agent

arXiv:2601.19932v2 Announce Type: replace Abstract: Coded language is an important part of human communication. It refers to cases where users intentionally encode meaning so that the surface text dif

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model

DGX agent

arXiv:2604.20806v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have made substantial advances in reasoning tasks at the Olympiad level. Nevertheless, current Olympiad-level mul

model-releasesarxiv-cs-ai
23 Apr 2026
← Previous
1…400401402403404…469
Next →