AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Psychologically-Grounded Graph Modeling for Interpretable Depression Detection

DGX agent

arXiv:2604.24126v1 Announce Type: new Abstract: Automatic depression detection from conversational interactions holds significant promise for scalable screening but remains hindered by severe data sca

model-releasesarxiv-cs-cl
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity

DGX agent

arXiv:2604.23972v1 Announce Type: cross Abstract: Knowledge graphs (KGs) are increasingly used to support large lan guage model (LLM) reasoning, but standard triplet-based KGs treat each relation as g

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Rewarding the Scientific Process: Process-Level Reward Modeling for Agentic Data Analysis

DGX agent

arXiv:2604.24198v1 Announce Type: cross Abstract: Process Reward Models (PRMs) have achieved remarkable success in augmenting the reasoning capabilities of Large Language Models (LLMs) within static d

safetyarxiv-cs-ai
28 Apr 2026
Research

Scoring, Reasoning, and Selecting the Best! Ensembling Large Language Models via a Peer-Review Process

DGX agent

arXiv:2512.23213v3 Announce Type: replace-cross Abstract: We propose LLM-PeerReview, an unsupervised LLM Ensemble method that selects the most ideal response from multiple LLM-generated candidates for

researcharxiv-cs-ai
28 Apr 2026
Hardware

Self-Rewarding Vision-Language Model via Reasoning Decomposition

DGX agent

arXiv:2508.19652v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) often suffer from visual hallucinations: generating things that are not consistent with visual inputs and language sho

hardwarearxiv-cs-cv
28 Apr 2026
Model Releases

The Rise of Large Language Models and the Direction and Impact of US Federal Research Funding

DGX agent

arXiv:2601.15485v2 Announce Type: replace-cross Abstract: Federal research funding shapes the direction, diversity, and impact of the US scientific enterprise. Large language models (LLMs) are rapidly

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models

DGX agent

arXiv:2512.10362v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) demonstrate impressive reasoning capabilities, but often fail to perceive fine-grained visual details

researcharxiv-cs-ai
28 Apr 2026
Research

Focus Session: Hardware and Software Techniques for Accelerating Multimodal Foundation Models

DGX agent

arXiv:2604.21952v1 Announce Type: cross Abstract: This work presents a multi-layered methodology for efficiently accelerating multimodal foundation models (MFMs). It combines hardware and software co-

researcharxiv-cs-ai
27 Apr 2026
Safety

How Large Language Models Balance Internal Knowledge with User and Document Assertions

DGX agent

arXiv:2604.22193v1 Announce Type: new Abstract: Large language models (LLMs) often need to balance their internal parametric knowledge with external information, such as user beliefs and content from

safetyarxiv-cs-cl
27 Apr 2026
Safety

Identifying and typifying demographic unfairness in phoneme-level embeddings of self-supervised speech recognition models

DGX agent

arXiv:2604.22631v1 Announce Type: new Abstract: Modern automatic speech recognition (ASR) systems have been observed to function better for certain speaker groups (SGs) than others, despite recent gai

safetyarxiv-cs-cl
27 Apr 2026
Research

Introducing Background Temperature to Characterise Hidden Randomness in Large Language Models

DGX agent

arXiv:2604.22411v1 Announce Type: new Abstract: Even when decoding with temperature T=0, large language models (LLMs) can produce divergent outputs for identical inputs. Recent work by Thinking Machin

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Sovereign Agentic Loops: Decoupling AI Reasoning from Execution in Real-World Systems

DGX agent

arXiv:2604.22136v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly issue API calls that mutate real systems, yet many current architectures pass stochastic model outputs

model-releasesarxiv-cs-lg
27 Apr 2026
Safety

System-Mediated Attention Imbalances Make Vision-Language Models Say Yes

DGX agent

arXiv:2601.12430v2 Announce Type: replace Abstract: Vision-language model (VLM) hallucination is commonly linked to imbalanced allocation of attention across input modalities: system, image and text.

safetyarxiv-cs-cl
27 Apr 2026
Agents

Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models

DGX agent

arXiv:2604.20994v1 Announce Type: cross Abstract: The growth of agentic AI has drawn significant attention to function calling Large Language Models (LLMs), which are designed to extend the capabiliti

agentsarxiv-cs-ai
24 Apr 2026
Applications

Capabilities and Evaluation Biases of Large Language Models in Classical Chinese Poetry Generation: A Case Study on Tang Poetry

DGX agent

arXiv:2510.15313v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly applied to creative domains, yet their performance in classical Chinese poetry generation and evaluati

applicationsarxiv-cs-cl
24 Apr 2026
Safety

Han Dan Xue Bu (Mimicry) or Qing Chu Yu Lan (Mastery)? A Cognitive Perspective on Reasoning Distillation in Large Language Models

DGX agent

arXiv:2601.05019v2 Announce Type: replace-cross Abstract: Recent Large Reasoning Models trained via reinforcement learning exhibit a 'natural' alignment with human cognitive costs. However, we show th

safetyarxiv-cs-ai
24 Apr 2026
Research

InVitroVision: a Multi-Modal AI Model for Automated Description of Embryo Development using Natural Language

DGX agent

arXiv:2604.21061v1 Announce Type: new Abstract: The application of artificial intelligence (AI) in IVF has shown promise in improving consistency and standardization of decisions, but often relies on

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Musical Score Understanding Benchmark: Evaluating Large Language Models' Comprehension of Complete Musical Scores

DGX agent

arXiv:2511.20697v4 Announce Type: replace-cross Abstract: Understanding complete musical scores entails integrated reasoning over pitch, rhythm, harmony, and large-scale structure, yet the ability of

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Who Defines Fairness? Target-Based Prompting for Demographic Representation in Generative Models

DGX agent

arXiv:2604.21036v1 Announce Type: new Abstract: Text-to-image(T2I) models like Stable Diffusion and DALL-E have made generative AI widely accessible, yet recent studies reveal that these systems often

safetyarxiv-cs-ai
24 Apr 2026
Research

Accumulated Aggregated D-Optimal Designs for Estimating Main Effects in Black-Box Models

DGX agent

arXiv:2510.08465v2 Announce Type: replace-cross Abstract: Estimating how individual input variables affect the output of a black-box model is a central task in explainable machine learning. However, e

researcharxiv-cs-lg
23 Apr 2026
Applications

An explicit operator explains end-to-end computation in the modern neural networks used for sequence and language modeling

DGX agent

arXiv:2604.20595v1 Announce Type: cross Abstract: We establish a mathematical correspondence between state space models, a state-of-the-art architecture for capturing long-range dependencies in data,

applicationsarxiv-cs-lg
23 Apr 2026
Research

Automated Description Generation of Cytologic Findings for Lung Cytological Images Using a Pretrained Vision Model and Dual Text Decoders: Preliminary Study

DGX agent

arXiv:2403.18151v2 Announce Type: replace-cross Abstract: Objective: Cytology plays a crucial role in lung cancer diagnosis. Pulmonary cytology involves cell morphological characterization in the spec

researcharxiv-cs-cv
23 Apr 2026
Model Releases

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models

DGX agent

arXiv:2604.16902v2 Announce Type: replace Abstract: Native Omni-modal Large Language Models (OLLMs) have shifted from pipeline architectures to unified representation spaces. However, this native inte

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Cold-Start Forecasting of New Product Life-Cycles via Conditional Diffusion Models

DGX agent

arXiv:2604.20370v1 Announce Type: new Abstract: Forecasting the life-cycle trajectory of a newly launched product is important for launch planning, resource allocation, and early risk assessment. This

researcharxiv-cs-lg
23 Apr 2026
Research

Depression Risk Assessment in Social Media via Large Language Models

DGX agent

arXiv:2604.19887v1 Announce Type: cross Abstract: Depression is one of the most prevalent and debilitating mental health conditions worldwide, frequently underdiagnosed and undertreated. The prolifera

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Development and Preliminary Evaluation of a Domain-Specific Large Language Model for Tuberculosis Care in South Africa

DGX agent

arXiv:2604.19776v1 Announce Type: new Abstract: Tuberculosis (TB) is one of the world's deadliest infectious diseases, and in South Africa, it contributes a significant burden to the country's health

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models

DGX agent

arXiv:2511.06209v4 Announce Type: replace Abstract: LLMs can solve complex tasks by generating long, multi-step reasoning chains. Test-time scaling (TTS) can further improve LLM performance by samplin

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts

DGX agent

arXiv:2604.19835v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for scaling large language models: frontier models routinely decouple total parameters f

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Infection-Reasoner: A Compact Vision-Language Model for Wound Infection Classification with Evidence-Grounded Clinical Reasoning

DGX agent

arXiv:2604.19937v1 Announce Type: cross Abstract: Assessing chronic wound infection from photographs is challenging because visual appearance varies across wound etiologies, anatomical locations, and

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?

DGX agent

arXiv:2601.13240v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at general programming but struggle with domain-specific software development, necessitating domain special

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure

DGX agent

arXiv:2604.20652v1 Announce Type: new Abstract: Large language models trained on human feedback may suppress fraud warnings when investors arrive already persuaded of a fraudulent opportunity. We test

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Machine learning moment closure models for the radiative transfer equation IV: enforcing symmetrizable hyperbolicity in two dimensions

DGX agent

arXiv:2604.20143v1 Announce Type: cross Abstract: This is our fourth work in the series on machine learning (ML) moment closure models for the radiative transfer equation (RTE). In the first three pap

researcharxiv-cs-lg
23 Apr 2026
Research

MasconCube: Fast and Accurate Gravity Modeling with an Explicit Representation

DGX agent

arXiv:2509.08607v3 Announce Type: replace-cross Abstract: The geodesy of irregularly shaped small bodies presents fundamental challenges for gravitational field modeling, particularly as deep space ex

researcharxiv-cs-lg
23 Apr 2026
Research

Mechanistic Interpretability Tool for AI Weather Models

DGX agent

arXiv:2604.20467v1 Announce Type: cross Abstract: Artificial Intelligence (AI) weather models are improving rapidly, and their forecasts are already competitive with long-established traditional Numer

researcharxiv-cs-lg
23 Apr 2026
Model Releases

Mitigating Hallucinations in Large Vision-Language Models without Performance Degradation

DGX agent

arXiv:2604.20366v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) exhibit powerful generative capabilities but frequently produce hallucinations that compromise output reliability.

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Spatio-temporal modelling of electric vehicle charging demand

DGX agent

arXiv:2604.19841v1 Announce Type: cross Abstract: Accurate forecasting of electric vehicle (EV) charging demand is critical for grid management and infrastructure planning. Yet the field continues to

model-releasesarxiv-cs-lg
23 Apr 2026
Research

Stabilising Generative Models of Attitude Change

DGX agent

arXiv:2604.19791v1 Announce Type: new Abstract: Attitude change - the process by which individuals revise their evaluative stances - has been explained by a set of influential but competing verbal the

researcharxiv-cs-ai
23 Apr 2026
Applications

Synthetic Flight Data Generation Using Generative Models

DGX agent

arXiv:2604.20293v1 Announce Type: new Abstract: The increasing adoption of synthetic data in aviation research offers a promising solution to data scarcity and confidentiality challenges. This study i

applicationsarxiv-cs-lg
23 Apr 2026
Research

Text Steganography with Dynamic Codebook and Multimodal Large Language Model

DGX agent

arXiv:2604.20269v1 Announce Type: cross Abstract: With the popularity of the large language models (LLMs), text steganography has achieved remarkable performance. However, existing methods still have

researcharxiv-cs-ai
23 Apr 2026
Model Releases

The GaoYao Benchmark: A Comprehensive Framework for Evaluating Multilingual and Multicultural Abilities of Large Language Models

DGX agent

arXiv:2604.20225v1 Announce Type: new Abstract: Evaluating the multilingual and multicultural capabilities of Large Language Models (LLMs) is essential for their global utility. However, current bench

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

The Ratchet Effect in Silico through Interaction-Driven Cumulative Intelligence in Large Language Models

DGX agent

arXiv:2507.21166v2 Announce Type: replace-cross Abstract: Human intelligence scales through cumulative cultural evolution (CCE), a ratchet process in which innovations are retained against entropic dr

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Toward Safe Autonomous Robotic Endovascular Interventions using World Models

DGX agent

arXiv:2604.20151v1 Announce Type: cross Abstract: Autonomous mechanical thrombectomy (MT) presents substantial challenges due to highly variable vascular geometries and the requirements for accurate,

model-releasesarxiv-cs-lg
23 Apr 2026
Applications

Towards reconstructing experimental sparse-view X-ray CT data with diffusion models

DGX agent

arXiv:2602.12755v2 Announce Type: replace Abstract: Diffusion-based image generators are promising priors for ill-posed inverse problems like sparse-view X-ray Computed Tomography (CT). As most studie

applicationsarxiv-cs-cv
23 Apr 2026
Research

WISCA: A Lightweight Model Transition Method to Improve LLM Training via Weight Scaling

DGX agent

arXiv:2508.16676v2 Announce Type: replace-cross Abstract: Transformer architecture gradually dominates the LLM field. Recent advances in training optimization for Transformer-based large language mode

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models

DGX agent

arXiv:2604.18612v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in complex reasoning tasks, while recent prompting strategies such as Chain-of-Thou

model-releasesarxiv-cs-ai
22 Apr 2026
Applications

Bayesian Event-Based Model for Disease Subtype and Stage Inference

DGX agent

arXiv:2512.03467v2 Announce Type: replace Abstract: Chronic diseases often progress differently across patients. Rather than randomly varying, there are typically a small number of subtypes for how a

applicationsarxiv-cs-lg
22 Apr 2026
Safety

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models

DGX agent

arXiv:2509.26238v4 Announce Type: replace Abstract: Monitoring large language models' (LLMs) activations is an effective way to detect harmful requests before they lead to unsafe outputs. However, tra

safetyarxiv-cs-lg
22 Apr 2026
Applications

Bridging the High-Frequency Data Gap: A Millisecond-Resolution Network Dataset for Advancing Time Series Foundation Models

DGX agent

arXiv:2603.16497v2 Announce Type: replace-cross Abstract: Time series foundation models (TSFMs) require diverse, real-world datasets to adapt across varying domains and temporal frequencies. However,

applicationsarxiv-cs-ai
22 Apr 2026
← Previous
1…116117118119120…1030
Next →