AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation

DGX agent

arXiv:2604.11424v1 Announce Type: new Abstract: Speech Language Models (SLMs) exhibit strong semantic understanding, yet their generated speech often sounds flat and fails to convey expressive intent,

model-releasesarxiv-cs-cl
14 Apr 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Bringing people together at AI for the Economy Forum

DGX agent

Google hosted its inaugural AI for the Economy Forum in Washington, D.C., co-hosted with MIT FutureTech, bringing together economists, industry leaders, policymakers, and experts to discuss AI's impac

model-releasesgoogle-ai
14 Apr 2026
Model Releases

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

DGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Camyla: Scaling Autonomous Research in Medical Image Segmentation

DGX agent

arXiv:2604.10696v1 Announce Type: new Abstract: We present Camyla, a system for fully autonomous research within the scientific domain of medical image segmentation. Camyla transforms raw datasets int

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Can Large Language Models Infer Causal Relationships from Real-World Text?

DGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?

DGX agent

arXiv:2511.21998v2 Announce Type: replace Abstract: Multi-modal Large Language Models (LLM) have advanced conversational abilities but struggle with providing live, interactive step-by-step guidance,

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

DGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration

DGX agent

arXiv:2509.17458v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models, such as Stable Diffusion, can produce high-quality and diverse images but often fail to achieve compositional

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation

DGX agent

arXiv:2604.10504v1 Announce Type: new Abstract: Current large language models (LLMs), even those explicitly trained for reasoning, often struggle with ambiguous content moderation cases due to mislead

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

DGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

DGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care

DGX agent

arXiv:2509.00891v3 Announce Type: replace Abstract: Real-world adoption of closed-loop insulin delivery systems (CLIDS) in type 1 diabetes remains low, driven not by technical failure, but by diverse

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

DGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

DGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Chrome now lets you turn AI prompts into repeatable ‘Skills’

DGX agent

Google is launching a new Chrome workflow feature that allows you to reuse your favorite Gemini commands across multiple webpages. Any AI prompts can now be saved as 'Skills' in the Chrome desktop bro

model-releasesthe-verge-ai
14 Apr 2026
Model Releases

ClaimDB: A Fact Verification Benchmark over Large Structured Data

DGX agent

arXiv:2601.14698v2 Announce Type: replace Abstract: Real-world fact-checking often involves verifying claims grounded in structured data at scale. Despite substantial progress in fact-verification ben

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems

DGX agent

arXiv:2604.10305v1 Announce Type: cross Abstract: Cooperative perception allows connected vehicles and roadside infrastructure to share sensor observations, creating a fused scene representation beyon

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Claude Code Routines are here! In addition to a schedule, you can now trigger templated agents via GitHub event or API – with our infra & yo…

DGX agent

Claude Code Routines are here! In addition to a schedule, you can now trigger templated agents via GitHub event or API – with our infra & your MCP+repos They've changed how we do docs, backlog mainten

model-releasesthariq--x
14 Apr 2026
Model Releases

ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents

DGX agent

arXiv:2604.10352v1 Announce Type: new Abstract: Stateful tool-using LLM agents treat the context window as working memory, yet today's agent harnesses manage residency and durability as best-effort, c

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation

DGX agent

arXiv:2604.11801v1 Announce Type: new Abstract: With the recent progress of Large Language Models (LLMs), there is a growing interest in applying these models to solve complex and challenging problems

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CocoaBench: Evaluating Unified Digital Agents in the Wild

DGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification

DGX agent

arXiv:2604.01687v2 Announce Type: replace Abstract: Anthropic proposes the concept of skills for LLM agents to tackle multi-step professional tasks that simple tool invocations cannot address. A tool

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CoFusion: Multispectral and Hyperspectral Image Fusion via Spectral Coordinate Attention

DGX agent

arXiv:2604.10584v1 Announce Type: new Abstract: Multispectral and Hyperspectral Image Fusion (MHIF) aims to reconstruct high-resolution images by integrating low-resolution hyperspectral images (LRHSI

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Comparative Analysis of Large Language Models in Healthcare

DGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Competing with AI Scientists: Agent-Driven Approach to Astrophysics Research

DGX agent

arXiv:2604.09621v1 Announce Type: new Abstract: We present an agent-driven approach to the construction of parameter inference pipelines for scientific data analysis. Our method leverages a multi-agen

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

COMPOSITE-Stem

DGX agent

arXiv:2604.09836v1 Announce Type: new Abstract: AI agents hold growing promise for accelerating scientific discovery; yet, a lack of frontier evaluations hinders adoption into real workflows. Expert-w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Computational Lesions in Multilingual Language Models Separate Shared and Language-specific Brain Alignment

DGX agent

arXiv:2604.10627v1 Announce Type: cross Abstract: How the brain supports language across different languages is a basic question in neuroscience and a useful test for multilingual artificial intellige

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Concept Drift Guided LayerNorm Tuning for Efficient Multimodal Metaphor Identification

DGX agent

arXiv:2505.11237v3 Announce Type: replace-cross Abstract: Metaphorical imagination, the ability to connect seemingly unrelated concepts, is fundamental to human cognition and communication. While unde

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Conflicts Make Large Reasoning Models Vulnerable to Attacks

DGX agent

arXiv:2604.09750v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable performance across diverse domains, yet their decision-making under conflicting objectives rema

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Using a Large Language Model

DGX agent

arXiv:2604.11287v1 Announce Type: new Abstract: Background: Large language models (LLMs) have been explored as tools for generating personalized exercise prescriptions, yet the consistency of outputs

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Context-Aware Semantic Segmentation via Stage-Wise Attention

DGX agent

arXiv:2601.11310v2 Announce Type: replace Abstract: Semantic ultra-high-resolution (UHR) image segmentation is essential in remote sensing applications such as aerial mapping and environmental monitor

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language Models

DGX agent

arXiv:2502.11008v2 Announce Type: replace Abstract: Counterfactual reasoning is widely recognized as one of the most challenging and intricate aspects of causality in artificial intelligence. In this

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Counting to Four is still a Chore for VLMs

DGX agent

arXiv:2604.10039v1 Announce Type: new Abstract: Vision--language models (VLMs) have achieved impressive performance on complex multimodal reasoning tasks, yet they still fail on simple grounding skill

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing

DGX agent

arXiv:2506.18438v2 Announce Type: replace Abstract: Editing natural images using textual descriptions in text-to-image diffusion models remains a significant challenge, particularly in achieving consi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CricBench: A Multilingual Benchmark for Evaluating LLMs in Cricket Analytics

DGX agent

arXiv:2512.21877v3 Announce Type: replace-cross Abstract: Cricket is the second most popular sport worldwide, with billions of fans seeking advanced statistical insights unavailable through standard w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Cross-Validated Cross-Channel Self-Attention and Denoising for Automatic Modulation Classification

DGX agent

arXiv:2604.10054v1 Announce Type: new Abstract: This study addresses a key limitation in deep learning Automatic Modulation Classification (AMC) models, which perform well at high signal-to-noise rati

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Cybersecurity Looks Like Proof of Work Now

DGX agent

Cybersecurity Looks Like Proof of Work Now The UK's AI Safety Institute recently published Our evaluation of Claude Mythos Preview’s cyber capabilities, their own independent analysis of Claude Mythos

model-releasessimon-willison
14 Apr 2026
Model Releases

Data-Efficient Semantic Segmentation of 3D Point Clouds via Open-Vocabulary Image Segmentation-based Pseudo-Labeling

DGX agent

arXiv:2604.11007v1 Announce Type: new Abstract: Semantic segmentation of 3D point cloud scenes is a crucial task for various applications. In real-world scenarios, training segmentation models often f

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

datasette PR #2689: Replace token-based CSRF with Sec-Fetch-Site header protection

DGX agent

datasette PR #2689: Replace token-based CSRF with Sec-Fetch-Site header protection Datasette has long protected against CSRF attacks using CSRF tokens, implemented using my asgi-csrf Python library. T

model-releasessimon-willison
14 Apr 2026
Model Releases

DDO-RM for LLM Preference Optimization: A Minimal Held-Out Benchmark against DPO

DGX agent

arXiv:2604.11119v1 Announce Type: cross Abstract: This paper reorganizes the current manuscript around the DPO versus DDO-RM preference-optimization project and focuses on two parts: the algorithmic v

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Dead Cognitions: A Census of Misattributed Insights

DGX agent

arXiv:2604.10288v1 Announce Type: new Abstract: This essay identifies a failure mode of AI chat systems that we term attribution laundering: the model performs substantive cognitive work and then rhet

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

DecepGPT: Schema-Driven Deception Detection with Multicultural Datasets and Robust Multimodal Learning

DGX agent

arXiv:2603.23916v2 Announce Type: replace-cross Abstract: Multimodal deception detection aims to identify deceptive behavior by analyzing audiovisual cues for forensics and security. In these high-sta

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Decompose, Mix, Adapt: A Unified Framework for Parameter-Efficient Neural Network Recombination and Compression

DGX agent

arXiv:2603.27383v2 Announce Type: replace Abstract: Parameter Recombination (PR) methods aim to efficiently compose the weights of a neural network for applications like Parameter-Efficient FineTuning

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Decomposing and Reducing Hidden Measurement Error in LLM Evaluation Pipelines

DGX agent

arXiv:2604.11581v1 Announce Type: new Abstract: LLM evaluations drive which models get deployed, which safety standards get adopted, and which research conclusions get published. Yet these scores carr

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

🚀 deepagents 0.5 release 👉 Async subagents - kick off background tasks on any Agent Protocol backed server while you continue to interact …

DGX agent

🚀 deepagents 0.5 release 👉 Async subagents - kick off background tasks on any Agent Protocol backed server while you continue to interact with the main agent. Start multiple background tasks in parall

model-releasesharrison-chase--x
14 Apr 2026
Model Releases

DeepReviewer 2.0: A Traceable Agentic System for Auditable Scientific Peer Review

DGX agent

arXiv:2604.09590v1 Announce Type: new Abstract: Automated peer review is often framed as generating fluent critique, yet reviewers and area chairs need judgments they can audit: where a concern applie

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Degradation-Consistent Paired Training for Robust AI-Generated Image Detection

DGX agent

arXiv:2604.10102v1 Announce Type: cross Abstract: AI-generated image detectors suffer significant performance degradation under real-world image corruptions such as JPEG compression, Gaussian blur, an

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Delta Rectified Flow Sampling for Text-to-Image Editing

DGX agent

arXiv:2509.05342v3 Announce Type: replace Abstract: We propose Delta Rectified Flow Sampling (DRFS), a novel inversion-free, path-aware editing framework within rectified flow models for text-to-image

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…436437438439440…464
Next →