AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,507 results
Model Releases

Anthropic says it has fixed three causes of recent Claude Code quality issues: reduced default reasoning, a caching bug, and a system prompt to reduce verbosity (Anthropic)

DGX agent

Anthropic: Anthropic says it has fixed three causes of recent Claude Code quality issues: reduced default reasoning, a caching bug, and a system prompt to reduce verbosity — We traced recent reports o

model-releasestechmeme
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Anthropic’s Mythos breach was humiliating

DGX agent

Anthropic's tightly controlled rollout of Claude Mythos has taken an awkward turn. After spending weeks insisting the AI model is so capable at cybersecurity that it is too dangerous to release public

model-releasesthe-verge-ai
23 Apr 2026
Model Releases

API pricing will be 5 per 1 million input tokens and 30 per 1 million output tokens, with a 1 million context window. (Remember, you will …

DGX agent

OpenAI's API pricing structure charges 5 per 1 million input tokens and 30 per 1 million output tokens, with support for a 1 million token context window. This pricing model reflects the higher cost o

model-releasessam-altman--x
23 Apr 2026
Model Releases

Apple fixes a bug that stored notifications for deleted messages on iPhone and iPad, following a report that police used it to extract deleted Signal messages (Lorenzo Franceschi-Bicchierai/TechCrunch)

DGX agent

Lorenzo Franceschi-Bicchierai / TechCrunch: Apple fixes a bug that stored notifications for deleted messages on iPhone and iPad, following a report that police used it to extract deleted Signal messag

model-releasestechmeme
23 Apr 2026
Model Releases

Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders

DGX agent

arXiv:2604.19974v1 Announce Type: cross Abstract: Large language models can be uncertain yet correct, or confident yet wrong, raising the question of whether their output-level uncertainty and their a

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Assessing the Robustness of Climate Foundation Models under No-Analog Distribution Shifts

DGX agent

arXiv:2603.23043v2 Announce Type: replace-cross Abstract: The accelerating pace of climate change introduces profound non-stationarities that challenge the ability of Machine Learning based climate em

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

ATIR: Towards Audio-Text Interleaved Contextual Retrieval

DGX agent

arXiv:2604.20267v1 Announce Type: cross Abstract: Audio carries richer information than text, including emotion, speaker traits, and environmental context, while also enabling lower-latency processing

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Automated Detection of Dosing Errors in Clinical Trial Narratives: A Multi-Modal Feature Engineering Approach with LightGBM

DGX agent

arXiv:2604.19759v1 Announce Type: new Abstract: Clinical trials require strict adherence to medication protocols, yet dosing errors remain a persistent challenge affecting patient safety and trial int

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Automatic Ontology Construction Using LLMs as an External Layer of Memory, Verification, and Planning for Hybrid Intelligent Systems

DGX agent

arXiv:2604.20795v1 Announce Type: new Abstract: This paper presents a hybrid architecture for intelligent systems in which large language models (LLMs) are extended with an external ontological memory

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Available on @ollama ! 🤝🤝

DGX agent

Available on @ollama ! 🤝🤝 Qwen 3.6 27B model is available on Ollama! Use it with all the integrations in Ollama or chat with the model. Chat with the model: ollama run qwen3.6:27b OpenClaw: ollama lau

model-releasesqwen--x
23 Apr 2026
Model Releases

AVISE: Framework for Evaluating the Security of AI Systems

DGX agent

arXiv:2604.20833v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across critical domains, their security vulnerabilities pose growing risks of high-p

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Benchmarking ResNet for Short-Term Hypoglycemia Classification with DiaData

DGX agent

arXiv:2511.02849v2 Announce Type: replace-cross Abstract: Individualized therapy is driven forward by medical data analysis, which provides insight into the patient's context. In particular, for Type

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Benefits of Low-Cost Bio-Inspiration in the Age of Overparametrization

DGX agent

arXiv:2604.20365v1 Announce Type: cross Abstract: While Central Pattern Generators (CPGs) and Multi-Layer Perceptrons (MLP) are widely used paradigms in robot control, few systematic studies have been

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Beyond Majority Voting: Towards Fine-grained and More Reliable Reward Signal for Test-Time Reinforcement Learning

DGX agent

arXiv:2512.15146v3 Announce Type: replace Abstract: Test-time reinforcement learning mitigates the reliance on annotated data by using majority voting results as pseudo-labels, emerging as a complemen

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models

DGX agent

arXiv:2604.16902v2 Announce Type: replace Abstract: Native Omni-modal Large Language Models (OLLMs) have shifted from pipeline architectures to unified representation spaces. However, this native inte

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Beyond the Crowd: LLM-Augmented Community Notes for Governing Health Misinformation

DGX agent

arXiv:2510.11423v3 Announce Type: replace-cross Abstract: Community Notes, the crowd-sourced misinformation governance system on X (formerly Twitter), allows users to flag misleading posts, attach con

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Bimanual Robot Manipulation via Multi-Agent In-Context Learning

DGX agent

arXiv:2604.20348v1 Announce Type: cross Abstract: Language Models (LLMs) have emerged as powerful reasoning engines for embodied control. In particular, In-Context Learning (ICL) enables off-the-shelf

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Bootstrapping Post-training Signals for Open-ended Tasks via Rubric-based Self-play on Pre-training Text

DGX agent

arXiv:2604.20051v1 Announce Type: new Abstract: Self-play has recently emerged as a promising paradigm to train Large Language Models (LLMs). In self-play, the target LLM creates the task input (e.g.,

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Braze launches agentic AI tools and Creative Studio, adds EU hosting for Decisioning Studio

DGX agent

Customer engagement platform company Braze Inc. today announced a new pair of agentic artificial intelligence tools for marketers and launched a new Creative Studio that links design software directly

model-releasessiliconangle
23 Apr 2026
Model Releases

Bridging Mechanistic Interpretability and Prompt Engineering with Gradient Ascent for Interpretable Persona Control

DGX agent

arXiv:2601.02896v2 Announce Type: replace Abstract: Controlling emergent behavioral personas (e.g., sycophancy, hallucination) in Large Language Models (LLMs) is critical for AI safety, yet remains a

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

btw in talking to friends the best framing for how to discuss GPT-Image-2-Thinking taking multiple tens of mins for generation and being abl…

DGX agent

btw in talking to friends the best framing for how to discuss GPT-Image-2-Thinking taking multiple tens of mins for generation and being able to oneshot QR codes and diagrams and logos and foods and f

model-releasesswyx--x
23 Apr 2026
Model Releases

Can 'AI' Be a Doctor? A Study of Empathy, Readability, and Alignment in Clinical LLMs

DGX agent

arXiv:2604.20791v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed in healthcare, yet their communicative alignment with clinical standards remains insufficiently

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Can We Locate and Prevent Stereotypes in LLMs?

DGX agent

arXiv:2604.19764v1 Announce Type: cross Abstract: Stereotypes in large language models (LLMs) can perpetuate harmful societal biases. Despite the widespread use of models, little is known about where

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence

DGX agent

arXiv:2603.28032v2 Announce Type: replace-cross Abstract: The convergence of low-altitude economies, embodied intelligence, and air-ground cooperative systems creates growing demand for simulation inf

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Catalyzing Informed Residential Energy Retrofit Decisions via Domain-Specific LLM

DGX agent

arXiv:2602.20181v2 Announce Type: replace-cross Abstract: Residential energy retrofit initiation is often stalled by an expertise gap, where homeowners lack the technical literacy required for structu

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs

DGX agent

arXiv:2604.20460v1 Announce Type: new Abstract: Safety-critical traffic reasoning requires contrastive consistency: models must detect true hazards when an accident occurs, and reliably reject plausib

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Chasing the Public Score: User Pressure and Evaluation Exploitation in Coding Agent Workflows

DGX agent

arXiv:2604.20200v1 Announce Type: new Abstract: Frontier coding agents are increasingly used in workflows where users supervise progress primarily through repeated improvement of a public score, namel

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Claude Code spend had gotten to $10.95M runrate peak at SemiAnalysis But then Opus 4.7 saved me. More token effecient for tasks, smarter, an…

DGX agent

Claude Code spend had gotten to $10.95M runrate peak at SemiAnalysis But then Opus 4.7 saved me. More token effecient for tasks, smarter, and no fast mode. Thank you @AnthropicAI You saved me from ban

model-releasesdylan-patel--x
23 Apr 2026
Model Releases

Claude is connecting directly to your personal apps like Spotify, Uber Eats, and TurboTax

DGX agent

Claude users can access more apps with Anthropic's AI now thanks to new connectors for everything from hiking to grocery shopping. Anthropic already supported connecting numerous work-related apps to

model-releasesthe-verge-ai
23 Apr 2026
Model Releases

CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values

DGX agent

arXiv:2509.03740v3 Announce Type: replace-cross Abstract: Vision-language models (VLMs) like CLIP have shown impressive zero-shot and few-shot learning capabilities across diverse applications. Howeve

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging

DGX agent

arXiv:2604.19750v1 Announce Type: cross Abstract: Recent advances in Large Language Model (LLM)-based agents have shown remarkable progress in code generation. However, current agent methods mainly re

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

DGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling

DGX agent

arXiv:2604.20720v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit performance disparities across languages, with naive multilingual fine-tuning frequently degrading performa

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

ConeSep: Cone-based Robust Noise-Unlearning Compositional Network for Composed Image Retrieval

DGX agent

arXiv:2604.20358v1 Announce Type: new Abstract: The Composed Image Retrieval (CIR) task provides a flexible retrieval paradigm via a reference image and modification text, but it heavily relies on exp

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Cooperative Profiles Predict Multi-Agent LLM Team Performance in AI for Science Workflows

DGX agent

arXiv:2604.20658v1 Announce Type: new Abstract: Multi-agent systems built from teams of large language models (LLMs) are increasingly deployed for collaborative scientific reasoning and problem-solvin

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

CRAFT: Training-Free Cascaded Retrieval for Tabular QA

DGX agent

arXiv:2505.14984v2 Announce Type: replace Abstract: Open-Domain Table Question Answering (TQA) involves retrieving relevant tables from a large corpus to answer natural language queries. Traditional d

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

CrowdStrike launches Project QuiltWorks coalition to tackle AI-discovered vulnerabilities

DGX agent

CrowdStrike Holdings Inc. today announced the launch of Project QuiltWorks, an industry coalition aimed at helping enterprises find and fix the wave of software vulnerabilities being surfaced by front

model-releasessiliconangle
23 Apr 2026
Model Releases

CyberCertBench: Evaluating LLMs in Cybersecurity Certification Knowledge

DGX agent

arXiv:2604.20389v1 Announce Type: cross Abstract: The rapid evolution and use of Large Language Models (LLMs) in professional workflows require an evaluation of their domain-specific knowledge against

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Databricks partners with OpenAI on GPT-5.5

DGX agent

Databricks has announced a partnership with OpenAI to integrate GPT-5.5, OpenAI's latest language model, into its data and AI platform. This collaboration enables Databricks customers to leverage GPT-

model-releasesdatabricks
23 Apr 2026
Model Releases

Day 0 vLLM support for Qwen3.6-27B! @vllm_project ♥️❤️

DGX agent

Day 0 vLLM support for Qwen3.6-27B! @vllm_project ♥️❤️ 🎉 Day-0 vLLM support for Qwen3.6-27B! Congrats to @Alibaba_Qwen on the new 27B dense model release. Looking forward to more of the Qwen3.6 series

model-releasesqwen--x
23 Apr 2026
Model Releases

Day 2 at Google Cloud Next: A marathon developer keynote

DGX agent

At Google Cloud, every day is Developer Day, but none so much as day 2 of Google Cloud Next, when we hold the developer keynote. This year’s topic? An in-depth look at Gemini Enterprise Agent Platform

model-releasesgoogle-cloud-ai
23 Apr 2026
Model Releases

Deepseek V4 on AI Gateway

DGX agent

Vercel announced support for DeepSeek V4 model through its AI Gateway service, enabling developers to access this AI model alongside other supported models on the platform. This addition expands Verce

model-releasesvercel-blog
23 Apr 2026
Model Releases

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

DGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories

DGX agent

arXiv:2604.20443v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been shown to possess Theory of Mind (ToM) abilities. However, it remains unclear whether this stems from robust rea

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Differentiable Conformal Training for LLM Reasoning Factuality

DGX agent

arXiv:2604.20098v1 Announce Type: new Abstract: Large Language Models (LLMs) frequently hallucinate, limiting their reliability in critical applications. Conformal Prediction (CP) addresses this by ca

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

DistortBench: Benchmarking Vision Language Models on Image Distortion Identification

DGX agent

arXiv:2604.19966v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in settings where sensitivity to low-level image degradations matters, including content moderatio

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Do Small Language Models Know When They're Wrong? Confidence-Based Cascade Scoring for Educational Assessment

DGX agent

arXiv:2604.19781v1 Announce Type: cross Abstract: Automated scoring of student work at scale requires balancing accuracy against cost and latency. In 'cascade' systems, small language models (LMs) han

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

'don't retweet this, don't retweet this, don't retweet this...' ah fuck it, life imitates art.

DGX agent

'don't retweet this, don't retweet this, don't retweet this...' ah fuck it, life imitates art. In Vending-Bench Arena (the multiplayer version of Vending-Bench with competition dynamics), GPT-5.5 actu

model-releasessam-altman--x
23 Apr 2026
← Previous
1…398399400401402…469
Next →