AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
Model Releases

Bidirectional Cross-Attention Fusion of High-Res RGB and Low-Res HSI for Multimodal Automated Waste Sorting

DGX agent

arXiv:2603.13941v2 Announce Type: replace Abstract: Growing waste streams and the transition to a circular economy require efficient automated waste sorting. In industrial settings, materials move on

model-releasesarxiv-cs-cv
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

BITS Pilani at SemEval-2026 Task 9: Structured Supervised Fine-Tuning with DPO Refinement for Polarization Detection

DGX agent

arXiv:2604.11121v1 Announce Type: new Abstract: The POLAR SemEval-2026 Shared Task aims to detect online polarization and focuses on the classification and identification of multilingual, multicultura

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

BLADE: Bayesian Langevin Active Discovery with Replica Exchange for Identification of Complex Systems

DGX agent

arXiv:2503.02983v2 Announce Type: replace-cross Abstract: Traditional methods for system discovery frequently struggle with efficient data usage and uncertainty quantification. Identifying the governi

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

BlasBench: An Open Benchmark for Irish Speech Recognition

DGX agent

arXiv:2604.10736v1 Announce Type: new Abstract: No open Irish-specific benchmark compares end-user ASR systems under a shared Irish-aware evaluation protocol. To solve this, we release BlasBench, an o

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection

DGX agent

arXiv:2604.10389v1 Announce Type: new Abstract: Terminology substitution errors in clinical notes, where one medical term is replaced by a linguistically valid but clinically different term, pose a pe

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

DGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Boxes2Pixels: Learning Defect Segmentation from Noisy SAM Masks

DGX agent

arXiv:2604.11162v1 Announce Type: new Abstract: Accurate defect segmentation is critical for industrial inspection, yet dense pixel-level annotations are rarely available. A common workaround is to co

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

BRIDGE and TCH-Net: Heterogeneous Benchmark and Multi-Branch Baseline for Cross-Domain IoT Botnet Detection

DGX agent

arXiv:2604.11324v1 Announce Type: cross Abstract: IoT botnet detection has advanced, yet most published systems are validated on a single dataset and rarely generalise across environments. Heterogeneo

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Bridging What the Model Thinks and How It Speaks: Self-Aware Speech Language Models for Expressive Speech Generation

DGX agent

arXiv:2604.11424v1 Announce Type: new Abstract: Speech Language Models (SLMs) exhibit strong semantic understanding, yet their generated speech often sounds flat and fails to convey expressive intent,

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Bringing people together at AI for the Economy Forum

DGX agent

Google hosted its inaugural AI for the Economy Forum in Washington, D.C., co-hosted with MIT FutureTech, bringing together economists, industry leaders, policymakers, and experts to discuss AI's impac

model-releasesgoogle-ai
14 Apr 2026
Model Releases

C-ReD: A Comprehensive Chinese Benchmark for AI-Generated Text Detection Derived from Real-World Prompts

DGX agent

arXiv:2604.11796v1 Announce Type: cross Abstract: Recently, large language models (LLMs) are capable of generating highly fluent textual content. While they offer significant convenience to humans, th

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Camyla: Scaling Autonomous Research in Medical Image Segmentation

DGX agent

arXiv:2604.10696v1 Announce Type: new Abstract: We present Camyla, a system for fully autonomous research within the scientific domain of medical image segmentation. Camyla transforms raw datasets int

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Can Large Language Models Infer Causal Relationships from Real-World Text?

DGX agent

arXiv:2505.18931v4 Announce Type: replace Abstract: Understanding and inferring causal relationships from texts is a core aspect of human cognition and is essential for advancing large language models

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?

DGX agent

arXiv:2511.21998v2 Announce Type: replace Abstract: Multi-modal Large Language Models (LLM) have advanced conversational abilities but struggle with providing live, interactive step-by-step guidance,

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CARE-ECG: Causal Agent-based Reasoning for Explainable and Counterfactual ECG Interpretation

DGX agent

arXiv:2604.10420v1 Announce Type: new Abstract: Large language models (LLMs) enable waveform-to-text ECG interpretation and interactive clinical questioning, yet most ECG-LLM systems still rely on wea

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration

DGX agent

arXiv:2509.17458v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models, such as Stable Diffusion, can produce high-quality and diverse images but often fail to achieve compositional

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CARO: Chain-of-Analogy Reasoning Optimization for Robust Content Moderation

DGX agent

arXiv:2604.10504v1 Announce Type: new Abstract: Current large language models (LLMs), even those explicitly trained for reasoning, often struggle with ambiguous content moderation cases due to mislead

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

DGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

DGX agent

arXiv:2503.21380v3 Announce Type: replace Abstract: The rapid advancement of large reasoning models has saturated existing math benchmarks, underscoring the urgent need for more challenging evaluation

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care

DGX agent

arXiv:2509.00891v3 Announce Type: replace Abstract: Real-world adoption of closed-loop insulin delivery systems (CLIDS) in type 1 diabetes remains low, driven not by technical failure, but by diverse

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

DGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ChemPro: A Progressive Chemistry Benchmark for Large Language Models

DGX agent

arXiv:2602.03108v3 Announce Type: replace Abstract: We introduce ChemPro, a progressive benchmark with 4100 natural language question-answer pairs in Chemistry, across 4 coherent sections of difficult

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Chrome now lets you turn AI prompts into repeatable ‘Skills’

DGX agent

Google is launching a new Chrome workflow feature that allows you to reuse your favorite Gemini commands across multiple webpages. Any AI prompts can now be saved as 'Skills' in the Chrome desktop bro

model-releasesthe-verge-ai
14 Apr 2026
Model Releases

ClaimDB: A Fact Verification Benchmark over Large Structured Data

DGX agent

arXiv:2601.14698v2 Announce Type: replace Abstract: Real-world fact-checking often involves verifying claims grounded in structured data at scale. Despite substantial progress in fact-verification ben

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Class-Adaptive Cooperative Perception for Multi-Class LiDAR-based 3D Object Detection in V2X Systems

DGX agent

arXiv:2604.10305v1 Announce Type: cross Abstract: Cooperative perception allows connected vehicles and roadside infrastructure to share sensor observations, creating a fused scene representation beyon

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Claude Code Routines are here! In addition to a schedule, you can now trigger templated agents via GitHub event or API – with our infra & yo…

DGX agent

Claude Code Routines are here! In addition to a schedule, you can now trigger templated agents via GitHub event or API – with our infra & your MCP+repos They've changed how we do docs, backlog mainten

model-releasesthariq--x
14 Apr 2026
Model Releases

ClawVM: Harness-Managed Virtual Memory for Stateful Tool-Using LLM Agents

DGX agent

arXiv:2604.10352v1 Announce Type: new Abstract: Stateful tool-using LLM agents treat the context window as working memory, yet today's agent harnesses manage residency and durability as best-effort, c

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation

DGX agent

arXiv:2604.11801v1 Announce Type: new Abstract: With the recent progress of Large Language Models (LLMs), there is a growing interest in applying these models to solve complex and challenging problems

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

CocoaBench: Evaluating Unified Digital Agents in the Wild

DGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification

DGX agent

arXiv:2604.01687v2 Announce Type: replace Abstract: Anthropic proposes the concept of skills for LLM agents to tackle multi-step professional tasks that simple tool invocations cannot address. A tool

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CoFusion: Multispectral and Hyperspectral Image Fusion via Spectral Coordinate Attention

DGX agent

arXiv:2604.10584v1 Announce Type: new Abstract: Multispectral and Hyperspectral Image Fusion (MHIF) aims to reconstruct high-resolution images by integrating low-resolution hyperspectral images (LRHSI

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Comparative Analysis of Large Language Models in Healthcare

DGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Competing with AI Scientists: Agent-Driven Approach to Astrophysics Research

DGX agent

arXiv:2604.09621v1 Announce Type: new Abstract: We present an agent-driven approach to the construction of parameter inference pipelines for scientific data analysis. Our method leverages a multi-agen

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

COMPOSITE-Stem

DGX agent

arXiv:2604.09836v1 Announce Type: new Abstract: AI agents hold growing promise for accelerating scientific discovery; yet, a lack of frontier evaluations hinders adoption into real workflows. Expert-w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Computational Lesions in Multilingual Language Models Separate Shared and Language-specific Brain Alignment

DGX agent

arXiv:2604.10627v1 Announce Type: cross Abstract: How the brain supports language across different languages is a basic question in neuroscience and a useful test for multilingual artificial intellige

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Concept Drift Guided LayerNorm Tuning for Efficient Multimodal Metaphor Identification

DGX agent

arXiv:2505.11237v3 Announce Type: replace-cross Abstract: Metaphorical imagination, the ability to connect seemingly unrelated concepts, is fundamental to human cognition and communication. While unde

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Conflicts Make Large Reasoning Models Vulnerable to Attacks

DGX agent

arXiv:2604.09750v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have achieved remarkable performance across diverse domains, yet their decision-making under conflicting objectives rema

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Using a Large Language Model

DGX agent

arXiv:2604.11287v1 Announce Type: new Abstract: Background: Large language models (LLMs) have been explored as tools for generating personalized exercise prescriptions, yet the consistency of outputs

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Context-Aware Semantic Segmentation via Stage-Wise Attention

DGX agent

arXiv:2601.11310v2 Announce Type: replace Abstract: Semantic ultra-high-resolution (UHR) image segmentation is essential in remote sensing applications such as aerial mapping and environmental monitor

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CounterBench: Evaluating and Improving Counterfactual Reasoning in Large Language Models

DGX agent

arXiv:2502.11008v2 Announce Type: replace Abstract: Counterfactual reasoning is widely recognized as one of the most challenging and intricate aspects of causality in artificial intelligence. In this

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Counting to Four is still a Chore for VLMs

DGX agent

arXiv:2604.10039v1 Announce Type: new Abstract: Vision--language models (VLMs) have achieved impressive performance on complex multimodal reasoning tasks, yet they still fail on simple grounding skill

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing

DGX agent

arXiv:2506.18438v2 Announce Type: replace Abstract: Editing natural images using textual descriptions in text-to-image diffusion models remains a significant challenge, particularly in achieving consi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CricBench: A Multilingual Benchmark for Evaluating LLMs in Cricket Analytics

DGX agent

arXiv:2512.21877v3 Announce Type: replace-cross Abstract: Cricket is the second most popular sport worldwide, with billions of fans seeking advanced statistical insights unavailable through standard w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Cross-Validated Cross-Channel Self-Attention and Denoising for Automatic Modulation Classification

DGX agent

arXiv:2604.10054v1 Announce Type: new Abstract: This study addresses a key limitation in deep learning Automatic Modulation Classification (AMC) models, which perform well at high signal-to-noise rati

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Cybersecurity Looks Like Proof of Work Now

DGX agent

Cybersecurity Looks Like Proof of Work Now The UK's AI Safety Institute recently published Our evaluation of Claude Mythos Preview’s cyber capabilities, their own independent analysis of Claude Mythos

model-releasessimon-willison
14 Apr 2026
Model Releases

Data-Efficient Semantic Segmentation of 3D Point Clouds via Open-Vocabulary Image Segmentation-based Pseudo-Labeling

DGX agent

arXiv:2604.11007v1 Announce Type: new Abstract: Semantic segmentation of 3D point cloud scenes is a crucial task for various applications. In real-world scenarios, training segmentation models often f

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

datasette PR #2689: Replace token-based CSRF with Sec-Fetch-Site header protection

DGX agent

datasette PR #2689: Replace token-based CSRF with Sec-Fetch-Site header protection Datasette has long protected against CSRF attacks using CSRF tokens, implemented using my asgi-csrf Python library. T

model-releasessimon-willison
14 Apr 2026
Model Releases

DDO-RM for LLM Preference Optimization: A Minimal Held-Out Benchmark against DPO

DGX agent

arXiv:2604.11119v1 Announce Type: cross Abstract: This paper reorganizes the current manuscript around the DPO versus DDO-RM preference-optimization project and focuses on two parts: the algorithmic v

model-releasesarxiv-cs-lg
14 Apr 2026
← Previous
1…437438439440441…465
Next →