AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
14 Apr 2026

Anthropic redesigns Claude Code on desktop, adding a sidebar for managing multiple sessions, a drag-and-drop layout, an integrated terminal, and a file editor (Claude)

Model ReleasesDGX agent

Claude: Anthropic redesigns Claude Code on desktop, adding a sidebar for managing multiple sessions, a drag-and-drop layout, an integrated terminal, and a file editor — Today, we're releasing a redesi

AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement

Model ReleasesDGX agent

arXiv:2604.10432v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies have emerged as a versatile paradigm for generalist robotic manipulation. However, precise object placement under

AOP-Smart: A RAG-Enhanced Large Language Model Framework for Adverse Outcome Pathway Analysis


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2604.10874v1 Announce Type: cross Abstract: Adverse Outcome Pathways (AOPs) are an important knowledge framework in toxicological research and risk assessment. In recent years, large language mo

Architecture-Agnostic Modality-Isolated Gated Fusion for Robust Multi-Modal Prostate MRI Segmentation

Model ReleasesDGX agent

arXiv:2604.10702v1 Announce Type: cross Abstract: Multi-parametric prostate MRI -- combining T2-weighted, apparent diffusion coefficient, and high b-value diffusion-weighted sequences -- is central to

Are We Recognizing the Jaguar or Its Background? A Diagnostic Framework for Jaguar Re-Identification

Model ReleasesDGX agent

arXiv:2604.09690v1 Announce Type: new Abstract: Jaguar re-identification (re-ID) from citizen-science imagery can look strong on standard retrieval metrics while still relying on the wrong evidence, s

Assessing the Pedagogical Readiness of Large Language Models as AI Tutors in Low-Resource Contexts: A Case Study of Nepal's K-10 Curriculum

Model ReleasesDGX agent

arXiv:2604.09619v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into educational ecosystems promises to democratize access to personalized tutoring, yet the readiness

ATANT v1.1: Positioning Continuity Evaluation Against Memory, Long-Context, and Agentic-Memory Benchmarks

Model ReleasesDGX agent

arXiv:2604.10981v1 Announce Type: new Abstract: ATANT v1.0 (arXiv:2604.06710) defined continuity as a system property with 7 required properties and introduced a 10-checkpoint, LLM-free evaluation met

AttnTrace: Contextual Attribution of Prompt Injection and Knowledge Corruption

Model ReleasesDGX agent

arXiv:2508.03793v2 Announce Type: replace Abstract: Long-context large language models (LLMs), such as Gemini-2.5-Pro and Claude-Sonnet-4, are increasingly used to empower advanced AI systems, includi

Audio Flamingo Next: Next-Generation Open Audio-Language Models for Speech, Sound, and Music

Model ReleasesDGX agent

arXiv:2604.10905v1 Announce Type: cross Abstract: We present Audio Flamingo Next (AF-Next), the next-generation and most capable large audio-language model in the Audio Flamingo series, designed to ad

Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing

Model ReleasesDGX agent

arXiv:2604.10708v1 Announce Type: cross Abstract: Recent progress in multimodal models has spurred rapid advances in audio understanding, generation, and editing. However, these capabilities are typic

AutoMS: Multi-Agent Evolutionary Search for Cross-Physics Inverse Microstructure Design

Model ReleasesDGX agent

arXiv:2603.27195v2 Announce Type: replace Abstract: Designing microstructures with coupled cross-physics objectives is a fundamental challenge where traditional topology optimization is often computat

AWS launches Amazon Bio Discovery, an AI-powered application designed to speed up drug development, giving scientists access to biological foundation models (Reuters)

Model ReleasesDGX agent

Reuters: AWS launches Amazon Bio Discovery, an AI-powered application designed to speed up drug development, giving scientists access to biological foundation models — Amazon's (AMZN.O) cloud unit on

Back to the Barn with LLAMAs: Evolving Pretrained LLM Backbones in Finetuning Vision Language Models

Model ReleasesDGX agent

arXiv:2604.10985v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have rapidly advanced by leveraging powerful pre-trained Large Language Models (LLMs) as core reasoning backbones. As new

BadGraph: A Backdoor Attack Against Latent Diffusion Model for Text-Guided Graph Generation

Model ReleasesDGX agent

arXiv:2510.20792v4 Announce Type: replace-cross Abstract: The rapid progress of graph generation has raised new security concerns, particularly regarding backdoor vulnerabilities. While prior work has

BankerToolBench: Evaluating AI Agents in End-to-End Investment Banking Workflows

Model ReleasesDGX agent

arXiv:2604.11304v1 Announce Type: new Abstract: Existing AI benchmarks lack the fidelity to assess economically meaningful progress on professional workflows. To evaluate frontier AI agents in a high-

BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs

Model ReleasesDGX agent

arXiv:2604.10528v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) demonstrate remarkable zero-shot recognition capabilities across a diverse spectrum of multimodal tasks, it yet rema

Beginner guide for anyone coming from ChatGPT who has never touched Claude before. No terminal, no tech talk. Ten steps, each with a plain explanation and a tip.

Model ReleasesDGX agent

This Reddit post from r/ChatGPT is a non-technical beginner's guide designed to help users transition from ChatGPT to Claude, structured as ten plain-language steps with practical tips — requiring no

Benchmarking Large Vision-Language Models on Fine-Grained Image Tasks: A Comprehensive Evaluation

Model ReleasesDGX agent

arXiv:2504.14988v3 Announce Type: replace Abstract: Recent advancements in Large Vision-Language Models (LVLMs) have demonstrated remarkable multimodal perception capabilities, garnering significant a

Benchmarking Vision-Language Models under Contradictory Virtual Content Attacks in Augmented Reality

Model ReleasesDGX agent

arXiv:2604.05510v2 Announce Type: replace Abstract: Augmented reality (AR) has rapidly expanded over the past decade. As AR becomes increasingly integrated into daily life, its security and reliabilit

Beyond Matching to Tiles: Bridging Unaligned Aerial and Satellite Views for Vision-Only UAV Navigation

Model ReleasesDGX agent

arXiv:2603.22153v3 Announce Type: replace-cross Abstract: Recent advances in cross-view geo-localization (CVGL) methods have shown strong potential for supporting unmanned aerial vehicle (UAV) navigat

Beyond Statistical Co-occurrence: Unlocking Intrinsic Semantics for Tabular Data Clustering

Model ReleasesDGX agent

arXiv:2604.10865v1 Announce Type: new Abstract: Deep Clustering (DC) has emerged as a powerful tool for tabular data analysis in real-world domains like finance and healthcare. However, most existing

Beyond the Beep: Scalable Collision Anticipation and Real-Time Explainability with BADAS-2.0

Model ReleasesDGX agent

arXiv:2604.05767v2 Announce Type: replace-cross Abstract: We present BADAS-2.0, the second generation of our collision anticipation system, building on BADAS-1.0, which showed that fine-tuning V-JEPA2

Bias Detection in Emergency Psychiatry: Linking Negative Language to Diagnostic Disparities

Model ReleasesDGX agent

arXiv:2509.02651v3 Announce Type: replace-cross Abstract: The emergency department (ED) is a high stress environment with increased risk of clinician bias exposure. In the United States, Black patient

BiCLIP: Domain Canonicalization via Structured Geometric Transformation

Model ReleasesDGX agent

arXiv:2603.08942v2 Announce Type: replace-cross Abstract: Recent advances in vision-language models (VLMs) have demonstrated remarkable zero-shot capabilities, yet adapting these models to specialized

Bidirectional Cross-Attention Fusion of High-Res RGB and Low-Res HSI for Multimodal Automated Waste Sorting

Model ReleasesDGX agent

arXiv:2603.13941v2 Announce Type: replace Abstract: Growing waste streams and the transition to a circular economy require efficient automated waste sorting. In industrial settings, materials move on

BITS Pilani at SemEval-2026 Task 9: Structured Supervised Fine-Tuning with DPO Refinement for Polarization Detection

Model ReleasesDGX agent

arXiv:2604.11121v1 Announce Type: new Abstract: The POLAR SemEval-2026 Shared Task aims to detect online polarization and focuses on the classification and identification of multilingual, multicultura

BLADE: Bayesian Langevin Active Discovery with Replica Exchange for Identification of Complex Systems

Model ReleasesDGX agent

arXiv:2503.02983v2 Announce Type: replace-cross Abstract: Traditional methods for system discovery frequently struggle with efficient data usage and uncertainty quantification. Identifying the governi

BlasBench: An Open Benchmark for Irish Speech Recognition

Model ReleasesDGX agent

arXiv:2604.10736v1 Announce Type: new Abstract: No open Irish-specific benchmark compares end-user ASR systems under a shared Irish-aware evaluation protocol. To solve this, we release BlasBench, an o

BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection

Model ReleasesDGX agent

arXiv:2604.10389v1 Announce Type: new Abstract: Terminology substitution errors in clinical notes, where one medical term is replaced by a linguistically valid but clinically different term, pose a pe

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

Model ReleasesDGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

Boxes2Pixels: Learning Defect Segmentation from Noisy SAM Masks

Model ReleasesDGX agent

arXiv:2604.11162v1 Announce Type: new Abstract: Accurate defect segmentation is critical for industrial inspection, yet dense pixel-level annotations are rarely available. A common workaround is to co

BRIDGE and TCH-Net: Heterogeneous Benchmark and Multi-Branch Baseline for Cross-Domain IoT Botnet Detection

Model ReleasesDGX agent

arXiv:2604.11324v1 Announce Type: cross Abstract: IoT botnet detection has advanced, yet most published systems are validated on a single dataset and rarely generalise across environments. Heterogeneo

Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation

Model ReleasesDGX agent

arXiv:2604.11424v1 Announce Type: new Abstract: Speech Language Models (SLMs) exhibit strong semantic understanding, yet their generated speech often sounds flat and fails to convey expressive intent,

Bringing people together at AI for the Economy Forum

Model ReleasesDGX agent

Google hosted its inaugural AI for the Economy Forum in Washington, D.C., co-hosted with MIT FutureTech, bringing together economists, industry leaders, policymakers, and experts to discuss AI's impac

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

Model ReleasesDGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

Camyla: Scaling Autonomous Research in Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.10696v1 Announce Type: new Abstract: We present Camyla, a system for fully autonomous research within the scientific domain of medical image segmentation. Camyla transforms raw datasets int

Can Large Language Models Infer Causal Relationships from Real-World Text?

Model ReleasesDGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?

Model ReleasesDGX agent

arXiv:2511.21998v2 Announce Type: replace Abstract: Multi-modal Large Language Models (LLM) have advanced conversational abilities but struggle with providing live, interactive step-by-step guidance,

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

Model ReleasesDGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration

Model ReleasesDGX agent

arXiv:2509.17458v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models, such as Stable Diffusion, can produce high-quality and diverse images but often fail to achieve compositional

CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation

Model ReleasesDGX agent

arXiv:2604.10504v1 Announce Type: new Abstract: Current large language models (LLMs), even those explicitly trained for reasoning, often struggle with ambiguous content moderation cases due to mislead

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

Model ReleasesDGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care

Model ReleasesDGX agent

arXiv:2509.00891v3 Announce Type: replace Abstract: Real-world adoption of closed-loop insulin delivery systems (CLIDS) in type 1 diabetes remains low, driven not by technical failure, but by diverse

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

Model ReleasesDGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

Chrome now lets you turn AI prompts into repeatable ‘Skills’

Model ReleasesDGX agent

Google is launching a new Chrome workflow feature that allows you to reuse your favorite Gemini commands across multiple webpages. Any AI prompts can now be saved as 'Skills' in the Chrome desktop bro

ClaimDB: A Fact Verification Benchmark over Large Structured Data

Model ReleasesDGX agent

arXiv:2601.14698v2 Announce Type: replace Abstract: Real-world fact-checking often involves verifying claims grounded in structured data at scale. Despite substantial progress in fact-verification ben

Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems

Model ReleasesDGX agent

arXiv:2604.10305v1 Announce Type: cross Abstract: Cooperative perception allows connected vehicles and roadside infrastructure to share sensor observations, creating a fused scene representation beyon

Claude Code Routines are here! In addition to a schedule, you can now trigger templated agents via GitHub event or API – with our infra & yo…

Model ReleasesDGX agent

Claude Code Routines are here! In addition to a schedule, you can now trigger templated agents via GitHub event or API – with our infra & your MCP+repos They've changed how we do docs, backlog mainten

ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2604.10352v1 Announce Type: new Abstract: Stateful tool-using LLM agents treat the context window as working memory, yet today's agent harnesses manage residency and durability as best-effort, c

CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation

Model ReleasesDGX agent

arXiv:2604.11801v1 Announce Type: new Abstract: With the recent progress of Large Language Models (LLMs), there is a growing interest in applying these models to solve complex and challenging problems

CocoaBench: Evaluating Unified Digital Agents in the Wild

Model ReleasesDGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification

Model ReleasesDGX agent

arXiv:2604.01687v2 Announce Type: replace Abstract: Anthropic proposes the concept of skills for LLM agents to tackle multi-step professional tasks that simple tool invocations cannot address. A tool

CoFusion: Multispectral and Hyperspectral Image Fusion via Spectral Coordinate Attention

Model ReleasesDGX agent

arXiv:2604.10584v1 Announce Type: new Abstract: Multispectral and Hyperspectral Image Fusion (MHIF) aims to reconstruct high-resolution images by integrating low-resolution hyperspectral images (LRHSI

Comparative Analysis of Large Language Models in Healthcare

Model ReleasesDGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

Competing with AI Scientists: Agent-Driven Approach to Astrophysics Research

Model ReleasesDGX agent

arXiv:2604.09621v1 Announce Type: new Abstract: We present an agent-driven approach to the construction of parameter inference pipelines for scientific data analysis. Our method leverages a multi-agen

COMPOSITE-Stem

Model ReleasesDGX agent

arXiv:2604.09836v1 Announce Type: new Abstract: AI agents hold growing promise for accelerating scientific discovery; yet, a lack of frontier evaluations hinders adoption into real workflows. Expert-w

Computational Lesions in Multilingual Language Models Separate Shared and Language-specific Brain Alignment

Model ReleasesDGX agent

arXiv:2604.10627v1 Announce Type: cross Abstract: How the brain supports language across different languages is a basic question in neuroscience and a useful test for multilingual artificial intellige

Concept Drift Guided LayerNorm Tuning for Efficient Multimodal Metaphor Identification

Model ReleasesDGX agent

arXiv:2505.11237v3 Announce Type: replace-cross Abstract: Metaphorical imagination, the ability to connect seemingly unrelated concepts, is fundamental to human cognition and communication. While unde

← Previous
1…349350351352353…372
Next →