AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,545 results
Model Releases

we're starting rollout of GPT-5.5-Cyber, a frontier cybersecurity model, to critical cyber defenders in the next few days. we will work with…

DGX agent

we're starting rollout of GPT-5.5-Cyber, a frontier cybersecurity model, to critical cyber defenders in the next few days. we will work with the entire ecosystem and the government to figure out trust

model-releasessam-altman--x
30 Apr 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today.

DGX agent

We've partnered with @OpenAI to offer GPT-5.5 in Devin at 50% off through May 14 starting today. GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible wi

model-releasescognition-ai--x
30 Apr 2026
Model Releases

We've partnered with @OpenAI to offer GPT-5.5 in Windsurf at 50% off through May 14 starting today.

DGX agent

Windsurf has announced a partnership with OpenAI to provide GPT-5.5 access within their platform at a 50% discount through May 14. This promotional offer began on the date of the announcement and appe

model-releaseswindsurf--x
30 Apr 2026
Model Releases

What Google Cloud announced in AI this month

DGX agent

Editor’s note: Want to keep up with the latest from Google Cloud? Check back here for a monthly recap of our latest updates, announcements, resources, events, learning opportunities, and more. We host

model-releasesgoogle-cloud-ai
30 Apr 2026
Model Releases

What people call 'distillation' is a super common practice (you use other models to benchmark your model, to evaluate your inputs or to add …

DGX agent

What people call 'distillation' is a super common practice (you use other models to benchmark your model, to evaluate your inputs or to add a little bit to your datasets) that in my opinion should be

model-releasesclem-delangue--x
30 Apr 2026
Model Releases

When to Retrieve During Reasoning: Adaptive Retrieval for Large Reasoning Models

DGX agent

arXiv:2604.26649v1 Announce Type: cross Abstract: Large reasoning models such as DeepSeek-R1 and OpenAI o1 generate extended chains of thought spanning thousands of tokens, yet their integration with

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Wordle 1,775 4/6 ⬛⬛🟨🟨⬛ 🟨⬛⬛⬛⬛ ⬛🟨🟨⬛🟩 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result where the player solved puzzle #1,775 in 4 attempts, using the color-coded feedback system (gray for incorrect letters, yellow for correct letters in wrong pos

model-releasesanthropic--x
30 Apr 2026
Model Releases

Work faster with Codex. https://chatgpt.com/codex/for-work/

DGX agent

Codex is OpenAI's AI tool designed to accelerate work productivity by generating and understanding code, enabling developers to write, debug, and complete programming tasks more efficiently. The resou

model-releasesopenai--x
30 Apr 2026
Model Releases

$3/million output tokens. Qwen 3.5 Plus is basically a frontier model. Let that sink in.

DGX agent

$3/million output tokens. Qwen 3.5 Plus is basically a frontier model. Let that sink in. Introducing Qwen3.6-Plus from @Alibaba_Qwen, a 1M-context model built for real-world agents, agentic coding, an

model-releasestogether-ai--x
29 Apr 2026
Model Releases

A Comparative Analysis on the Performance of Upper Confidence Bound Algorithms in Adaptive Deep Neural Networks

DGX agent

arXiv:2604.24810v1 Announce Type: new Abstract: Edge computing environments impose strict constraints on energy consumption and latency, making the deployment of deep neural networks a significant cha

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

A Comparative Study in Surgical AI: Datasets, Foundation Models, and Barriers to Med-AGI

DGX agent

arXiv:2603.27341v2 Announce Type: replace-cross Abstract: Recent Artificial Intelligence (AI) models have matched or exceeded human experts in several benchmarks of biomedical task performance, but su

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG perfo…

DGX agent

A million-token context window is not a strategy. 🛑 Our Head of DevRel, @RoieSchwabco , explains why dumping data is killing your RAG performance: 📍 One needle in a haystack? Easy. 📍 Multiple needles?

model-releasespinecone--x
29 Apr 2026
Model Releases

Adaptable phase retrieval for coherent transition radiation spectroscopy based on differentiable physics information

DGX agent

arXiv:2604.25489v1 Announce Type: cross Abstract: Coherent transition radiation (CTR) spectroscopy is a critical diagnostic for characterizing the longitudinal structure of relativistic electron bunch

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

AdaTooler-V: Adaptive Tool-Use for Images and Videos

DGX agent

arXiv:2512.16918v3 Announce Type: replace Abstract: Recent advances have shown that multimodal large language models (MLLMs) benefit from multimodal interleaved chain-of-thought (CoT) with vision tool

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

ADE: Adaptive Dictionary Embeddings -- Scaling Multi-Anchor Representations to Large Language Models

DGX agent

arXiv:2604.24940v1 Announce Type: new Abstract: Word embeddings are fundamental to natural language processing, yet traditional approaches represent each word with a single vector, creating representa

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Agent-Diff: Benchmarking LLM Agents on Enterprise API Tasks via Code Execution with State-Diff-Based Evaluation

DGX agent

arXiv:2602.11224v3 Announce Type: replace-cross Abstract: We present Agent-Diff, a novel benchmarking framework for evaluating agentic Large Language Models (LLMs) on real-world productivity software

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses

DGX agent

arXiv:2604.25850v1 Announce Type: new Abstract: Harnesses have become a central determinant of coding-agent performance, shaping how models interact with repositories, tools, and execution environment

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Align then Adapt: Rethinking Parameter-Efficient Transfer Learning in 4D Perception

DGX agent

arXiv:2602.23069v2 Announce Type: replace Abstract: Point cloud video understanding is critical for robotics as it accurately encodes motion and scene interaction. We recognize that 4D datasets are fa

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Am I the only one that still likes ChatGPT? And I use Claude also

DGX agent

A Reddit discussion from r/ChatGPT in which a user expresses their continued preference for ChatGPT while also using Claude, likely exploring whether other users share similar sentiments about ChatGPT

model-releasesr-chatgpt
29 Apr 2026
Model Releases

An Investigation of Linguistic Biases in LLM-Based Recommendations

DGX agent

arXiv:2604.25456v1 Announce Type: new Abstract: We investigate linguistic biases in LLM-based restaurant and product recommendations given prompts varying across Southern American English (AE), Indian

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Analyzing LLM Reasoning to Uncover Mental Health Stigma

DGX agent

arXiv:2604.25053v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly being explored for mental health applications, recent studies reveal that they can exhibit stigma to

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Application of a Mixture of Experts-based Foundation Model to the GlueX DIRC Detector

DGX agent

arXiv:2604.24775v1 Announce Type: cross Abstract: We present a Mixture-of-Experts-based foundation model applied to the GlueX DIRC detector at Jefferson Lab, demonstrating its utility as a unified fra

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

AQUA-Bench: Beyond Finding Answers to Knowing When There Are None in Audio Question Answering

DGX agent

arXiv:2601.12248v2 Announce Type: replace-cross Abstract: Recent advances in audio-aware large language models have shown strong performance on audio question answering. However, existing benchmarks m

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Architecture Determines Observability in Transformers

DGX agent

arXiv:2604.24801v1 Announce Type: new Abstract: Autoregressive transformers make confident errors, but activation monitoring can catch them only if the model preserves an internal signal that output c

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability require…

DGX agent

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability requires more than throughput, latency, and availability. It also r

model-releaseszhipu-ai--x
29 Apr 2026
Model Releases

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics

DGX agent

arXiv:2604.24916v1 Announce Type: new Abstract: We introduce asRoBallet, to the best of our knowledge, the first successful deployment of reinforcement learning (RL) on a humanoid ballbot hardware. Hi

model-releasesarxiv-cs-ro
29 Apr 2026
Model Releases

Auvik launches Aurora AI agents to speed ticket resolution and prevent outages

DGX agent

Information technology management software provider Auvik Networks Inc. today announced the launch of Auvik Aurora: artificial intelligence-powered IT agents that are designed to help IT professionals

model-releasessiliconangle
29 Apr 2026
Model Releases

Aviatrix launches AI agent containment platform for cloud workloads

DGX agent

Aviatrix Inc. today announced the launch of a new platform designed to contain artificial intelligence agents and enforce security controls and communications across AI workloads without changing AI a

model-releasessiliconangle
29 Apr 2026
Model Releases

Below-Chance Blindness: Prompted Underperformance in Small LLMs Produces Positional Bias Rather than Answer Avoidance

DGX agent

arXiv:2604.25249v1 Announce Type: new Abstract: Detecting sandbagging--the deliberate underperformance on capability evaluations--is an open problem in AI safety. We tested whether symptom validity te

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks

DGX agent

arXiv:2604.24955v1 Announce Type: new Abstract: As benchmarks grow in complexity, many apparent agent failures are not failures of the agent at all - they are failures of the benchmark itself: broken

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Benchmarking and Adapting On-Device LLMs for Clinical Decision Support

DGX agent

arXiv:2601.03266v2 Announce Type: replace Abstract: Large language models (LLMs) have rapidly advanced in clinical decision-making, yet the deployment of proprietary systems is hindered by privacy con

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Benchmarking and Improving GUI Agents in High-Dynamic Environments

DGX agent

arXiv:2604.25380v1 Announce Type: new Abstract: Recent advancements in Graphical User Interface (GUI) agents have predominantly focused on training paradigms like supervised fine-tuning (SFT) and rein

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Benchmarking Layout-Guided Diffusion Models through Unified Semantic-Spatial Evaluation in Closed and Open Settings

DGX agent

arXiv:2604.25358v1 Announce Type: new Abstract: Evaluating layout-guided text-to-image generative models requires assessing both semantic alignment with textual prompts and spatial fidelity to prescri

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization

DGX agent

arXiv:2604.25176v1 Announce Type: new Abstract: The digitization of multi-domain retail billing documents remains a challenging task due to variability in scan quality, layout heterogeneity, and domai

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal

DGX agent

arXiv:2509.09708v3 Announce Type: replace Abstract: Refusal on harmful prompts is a key safety behaviour in instruction-tuned large language models (LLMs), yet the internal causes of this behaviour re

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

BifDet: A 3D Bifurcation Detection Dataset for Airway-Tree Modeling

DGX agent

arXiv:2604.24999v1 Announce Type: new Abstract: Thoracic Computed Tomography (CT) scans offer detailed insights into the intricate branching network of the airway tree, which is essential for understa

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding

DGX agent

arXiv:2512.12087v3 Announce Type: replace Abstract: The growing demand for long-context inference capabilities in Large Language Models (LLMs) has intensified the computational and memory bottlenecks

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Building the compute infrastructure for the Intelligence Age

DGX agent

OpenAI outlines the computational infrastructure requirements and strategies necessary to support advanced AI systems in the emerging Intelligence Age. The piece likely discusses scaling challenges, h

model-releasesopenai
29 Apr 2026
Model Releases

CAN-QA: A Question-Answering Benchmark for Reasoning over In-Vehicle CAN Traffic

DGX agent

arXiv:2604.24935v1 Announce Type: cross Abstract: The Controller Area Network (CAN) is a safety-critical in-vehicle communication protocol that lacks built-in security mechanisms, making intrusion det

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

CGU-ILALab at FoodBench-QA 2026: Comparing Traditional and LLM-based Approaches for Recipe Nutrient Estimation

DGX agent

arXiv:2604.25774v1 Announce Type: new Abstract: Accurate nutrient estimation from unstructured recipe text is an important yet challenging problem in dietary monitoring, due to ambiguous ingredient te

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Cheaper, Better, Faster, Stronger: Robust Text-to-SQL without Chain-of-Thought or Fine-Tuning

DGX agent

arXiv:2505.14174v2 Announce Type: replace Abstract: LLMs are effective at code generation tasks like text-to-SQL, but is it worth the cost? Many state-of-the-art approaches use non-task-specific LLM t

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Citation Failure: Definition, Analysis and Efficient Mitigation

DGX agent

arXiv:2510.20303v3 Announce Type: replace Abstract: Citations from LLM-based RAG systems are supposed to simplify response verification. However, this goal is undermined in cases of citation failure,

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Codex can help you compare choices against your criteria and keep track of the tradeoffs.

DGX agent

Codex is an OpenAI tool that assists users in evaluating multiple options by comparing them against specified criteria and documenting the associated tradeoffs. This capability helps users make more i

model-releasesopenai--x
29 Apr 2026
Model Releases

Combating Visual Neglect and Semantic Drift in Large Multimodal Models for Enhanced Cross-Modal Retrieval

DGX agent

arXiv:2604.25273v1 Announce Type: new Abstract: Despite significant progress in Unified Multimodal Retrieval (UMR) powered by Large Multimodal Models (LMMs), existing embedding methods primarily focus

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Command Zero opens its autonomous security operations center platform with APIs and an MCP server

DGX agent

Cyber investigations platform provider Command Zero Inc. today released a set of application programming interface endpoints and a Model Context Protocol server for its autonomous security operations

model-releasessiliconangle
29 Apr 2026
Model Releases

Comparing Data Assimilation and Likelihood-Based Inference on Latent State Estimation in Agent-Based Models

DGX agent

arXiv:2509.17625v2 Announce Type: replace Abstract: In this paper, we present the first systematic comparison of Data Assimilation (DA) and Likelihood-Based Inference (LBI) in the context of an Agent-

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

Contrast-Enhanced Gating in GRUs for Robust Low-Data Sequence Learning

DGX agent

arXiv:2402.09034v3 Announce Type: replace Abstract: Activation functions govern how recurrent networks regulate and transmit information across temporal dependencies. Despite advances in sequence mode

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

CoRE: Concept-Reasoning Expansion for Continual Brain Lesion Segmentation

DGX agent

arXiv:2604.25376v1 Announce Type: new Abstract: Accurate brain lesion segmentation in MRI is vital for effective clinical diagnosis and treatment planning. Due to high annotation costs and strict data

model-releasesarxiv-cs-cv
29 Apr 2026
← Previous
1…374375376377378…470
Next →