AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,512 results
Model Releases

ControBench: An Interaction-Aware Benchmark for Controversial Discourse Analysis on Social Networks

DGX agent

arXiv:2605.00513v1 Announce Type: new Abstract: Understanding how people argue across ideological divides online is important for studying political polarization, misinformation, and content moderatio

model-releasesarxiv-cs-cl
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Disentangled Safety Adapters Enable Efficient Guardrails and Flexible Inference-Time Alignment

DGX agent

arXiv:2506.00166v2 Announce Type: replace-cross Abstract: Existing paradigms for ensuring AI safety, such as guardrail models and alignment training, often compromise either inference efficiency or de

safetyarxiv-cs-cl
4 May 2026
Model Releases

How Frontier LLMs Adapt to Neurodivergence Context: A Measurement Framework for Surface vs. Structural Change in System-Prompted Responses

DGX agent

arXiv:2605.00113v1 Announce Type: new Abstract: We examine if frontier chat-based large language models (LLMs) adjust their outputs based on neurodivergence (ND) context in system prompts and describe

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Make Your LVLM KV Cache More Lightweight

DGX agent

arXiv:2605.00789v1 Announce Type: new Abstract: Key-Value (KV) cache has become a de facto component of modern Large Vision-Language Models (LVLMs) for inference. While it enhances decoding efficiency

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Minimizing Human Intervention in Online Classification

DGX agent

arXiv:2510.23557v2 Announce Type: replace-cross Abstract: Training or fine-tuning large language model (LLM)-based systems often requires costly human feedback, yet there is limited understanding of h

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

PEACE: Cross-modal Enhanced Pediatric-Adult ECG Alignment for Robust Pediatric Diagnosis

DGX agent

arXiv:2605.00647v1 Announce Type: new Abstract: Automated pediatric electrocardiogram (ECG) diagnosis remains challenging because models trained predominantly on adult data suffer from substantial cro

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs

DGX agent

arXiv:2605.00814v1 Announce Type: new Abstract: While autoregressive Large Vision-Language Models (LVLMs) demonstrate remarkable proficiency in multimodal tasks, they face a 'Visual Signal Dilution' p

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Reasoning-Intensive Regression

DGX agent

arXiv:2508.21762v3 Announce Type: replace Abstract: AI researchers and practitioners increasingly apply large language models (LLMs) to what we call reasoning-intensive regression (RiR), i.e., deducin

model-releasesarxiv-cs-cl
4 May 2026
Research

RouteProfile: Elucidating the Design Space of LLM Profiles for Routing

DGX agent

arXiv:2605.00180v1 Announce Type: cross Abstract: As the large language model (LLM) ecosystem expands, individual models exhibit varying capabilities across queries, benchmarks, and domains, motivatin

researcharxiv-cs-cl
4 May 2026
Model Releases

Smart Profit-Aware Crop Advisory System: Kisan AI

DGX agent

arXiv:2605.00133v1 Announce Type: new Abstract: Modern crop advisory systems exhibit a critical limitation termed extit{economic blindness}. These systems primarily optimize for biological yield, ofte

model-releasesarxiv-cs-lg
4 May 2026
Safety

Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium

DGX agent

arXiv:2503.10990v2 Announce Type: replace-cross Abstract: Aligning large language models (LLMs) with diverse human preferences is critical for ensuring fairness and informed outcomes when deploying th

safetyarxiv-cs-lg
4 May 2026
Applications

The Power of Order: Fooling LLMs with Adversarial Table Permutations

DGX agent

arXiv:2605.00445v1 Announce Type: new Abstract: Large Language Models have achieved remarkable success and are increasingly deployed in critical applications involving tabular data, such as Table Ques

applicationsarxiv-cs-lg
4 May 2026
Model Releases

Why Do LLMs Struggle in Strategic Play? Broken Links Between Observations, Beliefs, and Actions

DGX agent

arXiv:2605.00226v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly tasked with strategic decision-making under incomplete information, such as in negotiation and policymakin

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

BatteryPass-12K: The First Dataset for the Novel Digital Battery Passport Conformance Task

DGX agent

arXiv:2604.26986v1 Announce Type: new Abstract: We introduce a novel task of digital battery passport (DBP) conformance classification and introduce the first public benchmark for the task: BatteryPas

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows

DGX agent

arXiv:2604.28139v1 Announce Type: cross Abstract: LLM agents are expected to complete end-to-end units of work across software tools, business services, and local workspaces. Yet many agent benchmarks

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures

DGX agent

arXiv:2604.28118v1 Announce Type: cross Abstract: Transformer models are widely deployed in critical AI applications, yet faults in their attention mechanisms, projections, and other internal componen

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Exploring Interaction Paradigms for LLM Agents in Scientific Visualization

DGX agent

arXiv:2604.27996v1 Announce Type: new Abstract: This paper examines how different types of large language model (LLM) agents perform on scientific visualization (SciVis) tasks, where users generate vi

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Test-taking to Cognitive Scaffolding: A Pedagogical Diagnostic Benchmark for LLMs on English Standardized Tests

DGX agent

arXiv:2505.17056v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly integrated into educational tools, current evaluations on standardized tests predominantly fo

model-releasesarxiv-cs-ai
1 May 2026
Research

Generative Human Geometry Distribution

DGX agent

arXiv:2503.01448v5 Announce Type: replace Abstract: Realistic human geometry generation is an important yet challenging task, requiring both the preservation of fine clothing details and the accurate

researcharxiv-cs-cv
1 May 2026
Model Releases

GlowQ: Group-Shared LOw-Rank Approximation for Quantized LLMs

DGX agent

arXiv:2603.25385v2 Announce Type: replace-cross Abstract: Quantization techniques such as BitsAndBytes, AWQ, and GPTQ are widely used as a standard method in deploying large language models but often

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions

DGX agent

arXiv:2604.27763v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) offers a transformative interface for Web3, yet existing benchmarks fail to capture the complexity of tran

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

InteractWeb-Bench: Can Multimodal Agent Escape Blind Execution in Interactive Website Generation?

DGX agent

arXiv:2604.27419v1 Announce Type: new Abstract: With the advancement of multimodal large language models (MLLMs) and coding agents, the website development has shifted from manual programming to agent

model-releasesarxiv-cs-ai
1 May 2026
Applications

Making Logic a First-Class Citizen in Generative ML for Networking

DGX agent

arXiv:2506.23964v3 Announce Type: replace-cross Abstract: Generative ML models are increasingly popular in networking for tasks such as telemetry imputation, prediction, and synthetic trace generation

applicationsarxiv-cs-lg
1 May 2026
Model Releases

MCPHunt: An Evaluation Framework for Cross-Boundary Data Propagation in Multi-Server MCP Agents

DGX agent

arXiv:2604.27819v1 Announce Type: new Abstract: Multi-server MCP agents create an information-flow control problem: faithful tool composition can turn individually benign read/write permissions into c

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ORFS-agent: Tool-Using Agents for Chip Design Optimization

DGX agent

arXiv:2506.08332v3 Announce Type: replace Abstract: Machine learning has been widely used to optimize complex engineering workflows across numerous domains. In integrated circuit design, modern flows

model-releasesarxiv-cs-ai
1 May 2026
Applications

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation

DGX agent

arXiv:2604.27747v1 Announce Type: cross Abstract: Large language model (LLM)-based generative list-wise recommendation has advanced rapidly, but decoding remains sequential and thus latency-prone. To

applicationsarxiv-cs-ai
1 May 2026
Model Releases

Post-Optimization Adaptive Rank Allocation for LoRA

DGX agent

arXiv:2604.27796v1 Announce Type: new Abstract: Exponential growth in the scale of modern foundation models has led to the widespread adoption of Low-Rank Adaptation (LoRA) as a parameter-efficient fi

model-releasesarxiv-cs-ai
1 May 2026
Applications

Predicting Covariate-Driven Spatial Deformation for Nonstationary Gaussian Processes

DGX agent

arXiv:2604.27280v1 Announce Type: new Abstract: Nonstationary Gaussian processes (GPs) are essential for modeling complex, locally heterogeneous spatial data. A common modeling approach is the spatial

applicationsarxiv-cs-lg
1 May 2026
Model Releases

Reinforced Agent: Inference-Time Feedback for Tool-Calling Agents

DGX agent

arXiv:2604.27233v1 Announce Type: new Abstract: Tool-calling agents are evaluated on tool selection, parameter accuracy, and scope recognition, yet LLM trajectory assessments remain inherently post-ho

model-releasesarxiv-cs-ai
1 May 2026
Tutorials

Sequential Inference for Gaussian Processes: A Signal Processing Perspective

DGX agent

arXiv:2604.28163v1 Announce Type: cross Abstract: The proliferation of capable and efficient machine learning (ML) models marks one of the strongest methodological shifts in signal processing (SP) in

tutorialsarxiv-cs-lg
1 May 2026
Model Releases

Step-level Optimization for Efficient Computer-use Agents

DGX agent

arXiv:2604.27151v1 Announce Type: new Abstract: Computer-use agents provide a promising path toward general software automation because they can interact directly with arbitrary graphical user interfa

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

AdaMem: Adaptive User-Centric Memory for Long-Horizon Dialogue Agents

DGX agent

arXiv:2603.16496v2 Announce Type: replace Abstract: Large language model (LLM) agents increasingly rely on external memory to support long-horizon interaction, personalized assistance, and multi-step

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Associative-State Universal Transformers: Sparse Retrieval Meets Structured Recurrence

DGX agent

arXiv:2604.25930v1 Announce Type: new Abstract: We study whether a structured recurrent state can serve as a compact associative backbone for language modeling while still supporting exact retrieval.

model-releasesarxiv-cs-cl
30 Apr 2026
Research

ChartVerse: Scaling Chart Reasoning via Reliable Programmatic Synthesis from Scratch

DGX agent

arXiv:2601.13606v2 Announce Type: replace Abstract: Chart reasoning is a critical capability for Vision Language Models (VLMs). However, the development of open-source models is severely hindered by t

researcharxiv-cs-cv
30 Apr 2026
Model Releases

ClassEval-Pro: A Cross-Domain Benchmark for Class-Level Code Generation

DGX agent

arXiv:2604.26923v1 Announce Type: cross Abstract: LLMs have achieved strong results on both function-level code synthesis and repository-level code modification, yet a capability that falls between th

model-releasesarxiv-cs-cl
30 Apr 2026
Safety

EvoSelect: Data-Efficient LLM Evolution for Targeted Task Adaptation

DGX agent

arXiv:2604.26170v1 Announce Type: new Abstract: Adapting large language models (LLMs) to a targeted task efficiently and effectively remains a fundamental challenge. Such adaptation often requires ite

safetyarxiv-cs-cl
30 Apr 2026
Research

FaaSMoE: A Serverless Framework for Multi-Tenant Mixture-of-Experts Serving

DGX agent

arXiv:2604.26881v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models offer high capacity with efficient inference cost by activating a small subset of expert models per input. However, de

researcharxiv-cs-lg
30 Apr 2026
Model Releases

Human-in-the-Loop Benchmarking of Heterogeneous LLMs for Automated Competency Assessment in Secondary Level Mathematics

DGX agent

arXiv:2604.26607v1 Announce Type: new Abstract: As Competency-Based Education (CBE) is gaining traction around the world, the shift from marks-based assessment to qualitative competency mapping is a m

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

HumanOmni-Speaker: Identifying Who said What and When

DGX agent

arXiv:2603.21664v2 Announce Type: replace Abstract: While Omni-modal Large Language Models have made strides in joint sensory processing, they fundamentally struggle with a cornerstone of human intera

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

Learning Neural Operator Surrogates for the Black Hole Accretion Code

DGX agent

arXiv:2604.25985v1 Announce Type: cross Abstract: General-relativistic magnetohydrodynamic (GR-MHD) simulations are essential for studying black hole accretion, relativistic jets, and magnetic reconne

model-releasesarxiv-cs-lg
30 Apr 2026
Applications

Pointer-CAD: Unifying B-Rep and Command Sequences via Pointer-based Edges & Faces Selection

DGX agent

arXiv:2603.04337v2 Announce Type: replace-cross Abstract: Constructing computer-aided design (CAD) models is labor-intensive but essential for engineering and manufacturing. Recent advances in Large L

applicationsarxiv-cs-cl
30 Apr 2026
Local Ai

Privacy-Preserving Federated Learning Framework for Distributed Chemical Process Optimization

DGX agent

arXiv:2604.26073v1 Announce Type: cross Abstract: Industrial chemical plants often operate under strict data confidentiality constraints, making centralized data-driven process modeling difficult. Fed

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

Reasoning Gets Harder for LLMs Inside A Dialogue

DGX agent

arXiv:2603.20133v2 Announce Type: replace Abstract: Large Language Models (LLMs) achieve strong performance on many reasoning benchmarks, yet these evaluations typically focus on isolated tasks that d

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

SciMDR: Advancing Scientific Multimodal Document Reasoning

DGX agent

arXiv:2603.12249v2 Announce Type: replace-cross Abstract: Constructing scientific multimodal document reasoning datasets for foundation model training involves an inherent trade-off among scale, faith

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences

DGX agent

arXiv:2509.11295v2 Announce Type: replace Abstract: Developing effective prompts demands significant cognitive investment to generate reliable, high-quality responses from Large Language Models (LLMs)

model-releasesarxiv-cs-cl
30 Apr 2026
Research

Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall

DGX agent

arXiv:2505.13963v3 Announce Type: replace Abstract: Quantization methods are widely used to accelerate inference and streamline the deployment of large language models (LLMs). Although quantization's

researcharxiv-cs-cl
30 Apr 2026
Model Releases

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

DGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses

DGX agent

arXiv:2604.25850v1 Announce Type: new Abstract: Harnesses have become a central determinant of coding-agent performance, shaping how models interact with repositories, tools, and execution environment

model-releasesarxiv-cs-cl
29 Apr 2026
← Previous
1…375376377378379…1074
Next →