AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
22 Apr 2026

Beyond Marginal Distributions: A Framework to Evaluate the Representativeness of Demographic-Aligned LLMs

SafetyDGX agent

arXiv:2601.15755v3 Announce Type: replace Abstract: Large language models are increasingly used to represent human opinions, values, or beliefs, and their steerability towards these ideals is an activ

Beyond One Output: Visualizing and Comparing Distributions of Language Model Generations

ResearchDGX agent

arXiv:2604.18724v1 Announce Type: new Abstract: Users typically interact with and evaluate language models via single outputs, but each output is just one sample from a broad distribution of possible

Beyond Rating: A Comprehensive Evaluation and Benchmark for AI Reviews

Model ReleasesDGX agent

arXiv:2604.19502v1 Announce Type: new Abstract: The rapid adoption of Large Language Models (LLMs) has spurred interest in automated peer review; however, progress is currently stifled by benchmarks t

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications

SafetyDGX agent

arXiv:2604.19281v1 Announce Type: cross Abstract: The use of Large Language Models (LLMs) to support patients in addressing medical questions is becoming increasingly prevalent. However, most of the m

Beyond the Bellman Fixed Point: Geometry and Fast Policy Identification in Value Iteration

SafetyDGX agent

arXiv:2604.17457v2 Announce Type: replace-cross Abstract: Dynamic programming is one of the most fundamental methodologies for solving Markov decision problems. Among its many variants, Q-value iterat

Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development

AgentsDGX agent

arXiv:2604.16323v2 Announce Type: replace-cross Abstract: As autonomous coding agents become deeply embedded in software development workflows, their high operational velocity introduces a critical ov

Bootstrapping Code Translation with Weighted Multilanguage Exploration

ResearchDGX agent

arXiv:2601.03512v2 Announce Type: replace-cross Abstract: Code translation across multiple programming languages is essential yet challenging due to two vital obstacles: scarcity of parallel data pair

Breaking the Illusion: Consensus-Based Generative Mitigation of Adversarial Illusions in Multi-Modal Embeddings

SafetyDGX agent

arXiv:2511.21893v2 Announce Type: replace Abstract: Multi-modal foundation models align images, text, and other modalities in a shared embedding space but remain vulnerable to adversarial illusions [3

Bridging Foundation Models and ASTM Metallurgical Standards for Automated Grain Size Estimation from Microscopy Images

Model ReleasesDGX agent

arXiv:2604.18957v1 Announce Type: new Abstract: Extracting standardized metallurgical metrics from microscopy images remains challenging due to complex grain morphology and the data demands of supervi

Bridging Semantics and Geometry: A Decoupled LVLM-SAM Framework for Reasoning Segmentation in Optical Remote Sensing

SafetyDGX agent

arXiv:2512.19302v2 Announce Type: replace Abstract: Large Vision--Language Models (LVLMs) hold great promise for advancing optical remote sensing (RS) analysis, yet existing reasoning segmentation fra

Bridging the High-Frequency Data Gap: A Millisecond-Resolution Network Dataset for Advancing Time Series Foundation Models

ApplicationsDGX agent

arXiv:2603.16497v2 Announce Type: replace-cross Abstract: Time series foundation models (TSFMs) require diverse, real-world datasets to adapt across varying domains and temporal frequencies. However,

Budgeted Online Influence Maximization

ApplicationsDGX agent

arXiv:2604.19672v1 Announce Type: new Abstract: We introduce a new budgeted framework for online influence maximization, considering the total cost of an advertising campaign instead of the common car

Byzantine-tolerant distributed learning of finite mixture models

Model ReleasesDGX agent

arXiv:2407.13980v3 Announce Type: replace-cross Abstract: Traditional statistical methods need to be updated to work with modern distributed data storage paradigms. A common approach is the split-and-

CAHAL: Clinically Applicable resolution enHAncement for Low-resolution MRI scans

SafetyDGX agent

arXiv:2604.18781v1 Announce Type: new Abstract: Large-scale automated morphometric analysis of brain MRI is limited by the thick-slice, anisotropic acquisitions prevalent in routine clinical practice.

Calibrating Scientific Foundation Models with Inference-Time Stochastic Attention

Model ReleasesDGX agent

arXiv:2604.19530v1 Announce Type: new Abstract: Transformer-based scientific foundation models are increasingly deployed in high-stakes settings, but current architectures give deterministic outputs a

Can AI-Generated Persuasion Be Detected? Persuaficial Benchmark and AI vs. Human Linguistic Differences

Model ReleasesDGX agent

arXiv:2601.04925v2 Announce Type: replace Abstract: Large Language Models (LLMs) can generate highly persuasive text, raising concerns about their misuse for propaganda, manipulation, and other harmfu

Can Continual Pre-training Bridge the Performance Gap between General-purpose and Specialized Language Models in the Medical Domain?

Model ReleasesDGX agent

arXiv:2604.19394v1 Announce Type: new Abstract: This paper narrows the performance gap between small, specialized models and significantly larger general-purpose models through domain adaptation via c

Can We Build Scene Graphs, Not Classify Them? FlowSG: Progressive Image-Conditioned Scene Graph Generation with Flow Matching

Local AiDGX agent

arXiv:2604.18623v1 Announce Type: new Abstract: Scene Graph Generation (SGG) unifies object localization and visual relationship reasoning by predicting boxes and subject-predicate-object triples. Yet

Capturing Classic Authorial Style in Long-Form Story Generation with GRPO Fine-Tuning

SafetyDGX agent

arXiv:2512.05747v3 Announce Type: replace Abstract: Evaluating and optimising authorial style in long-form story generation remains challenging because style is often assessed with ad hoc prompting an

CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmark

Model ReleasesDGX agent

arXiv:2505.16968v4 Announce Type: replace-cross Abstract: Cross-architecture GPU code transpilation is essential for unlocking low-level hardware portability, yet no scalable solution exists. We intro

CAST: Modeling Semantic-Level Transitions for Complementary-Aware Sequential Recommendation

Model ReleasesDGX agent

arXiv:2604.19414v1 Announce Type: cross Abstract: Sequential Recommendation (SR) aims to predict the next interaction of a user based on their behavior sequence, where complementary relations often pr

Cell-Based Representation of Relational Binding in Language Models

ResearchDGX agent

arXiv:2604.19052v1 Announce Type: new Abstract: Understanding a discourse requires tracking entities and the relations that hold between them. While Large Language Models (LLMs) perform well on relati

CentaurTA Studio: A Self-Improving Human-Agent Collaboration System for Thematic Analysis

SafetyDGX agent

arXiv:2604.18589v1 Announce Type: cross Abstract: Thematic analysis is difficult to scale: manual workflows are labor-intensive, while fully automated pipelines often lack controllability and transpar

Centralized Copy-Paste: Enhanced Data Augmentation Strategy for Wildland Fire Semantic Segmentation

ResearchDGX agent

arXiv:2507.06321v2 Announce Type: replace Abstract: Collecting and annotating images for the purpose of training segmentation models is often cost prohibitive. In the domain of wildland fire science,

Chain-of-Thought as a Lens: Evaluating Structured Reasoning Alignment between Human Preferences and Large Language Models

SafetyDGX agent

arXiv:2511.06168v3 Announce Type: replace Abstract: This paper primarily demonstrates a method to quantitatively assess the alignment between multi-step, structured reasoning in large language models

Characterizing AlphaEarth Embedding Geometry for Agentic Environmental Reasoning

Model ReleasesDGX agent

arXiv:2604.18715v1 Announce Type: cross Abstract: Earth observation foundation models encode land surface information into dense embedding vectors, yet the geometric structure of these representations

Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Language

Model ReleasesDGX agent

arXiv:2604.19667v1 Announce Type: cross Abstract: At present, executable visual workflows have emerged as a mainstream paradigm in real-world industrial deployments, offering strong reliability and co

Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence

ResearchDGX agent

arXiv:2602.12851v3 Announce Type: replace-cross Abstract: Deploying expressive learning models directly on programmable dataplanes promises line-rate, low-latency traffic analysis but remains hindered

Choose Your Own Adventure: Non-Linear AI-Assisted Programming with EvoGraph

ResearchDGX agent

arXiv:2604.18883v1 Announce Type: cross Abstract: Current AI-assisted programming tools are predominantly linear and chat-based, which deviates from the iterative and branching nature of programming i

CityRAG: Stepping Into a City via Spatially-Grounded Video Generation

AgentsDGX agent

arXiv:2604.19741v1 Announce Type: new Abstract: We address the problem of generating a 3D-consistent, navigable environment that is spatially grounded: a simulation of a real location. Existing video

ClawNet: Human-Symbiotic Agent Network for Cross-User Autonomous Cooperation

AgentsDGX agent

arXiv:2604.19211v1 Announce Type: new Abstract: Current AI agent frameworks have made remarkable progress in automating individual tasks, yet all existing systems serve a single user. Human productivi

CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation

Model ReleasesDGX agent

arXiv:2602.20409v2 Announce Type: replace Abstract: Recent vision-language models (VLMs) such as CLIP demonstrate impressive cross-modal reasoning, extending beyond images to 3D perception. Yet, these

Cloning Deterministic Worlds: The Critical Role of Latent Geometry in Long-Horizon World Models

SafetyDGX agent

arXiv:2510.26782v3 Announce Type: replace-cross Abstract: A world model is an internal model that simulates how the world evolves. Given past observations and actions, it predicts the future physical

Co-Refine: AI-Powered Tool Supporting Qualitative Analysis

ResearchDGX agent

arXiv:2604.19309v1 Announce Type: cross Abstract: Qualitative coding relies on a researcher's application of codes to textual data. As coding proceeds across large datasets, interpretations of codes o

CoCo-SAM3: Harnessing Concept Conflict in Open-Vocabulary Semantic Segmentation

ResearchDGX agent

arXiv:2604.19648v1 Announce Type: cross Abstract: SAM3 advances open-vocabulary semantic segmentation by introducing a prompt-driven mask generation paradigm. However, in multi-class open-vocabulary s

CoDA: Towards Effective Cross-domain Knowledge Transfer via CoT-guided Domain Adaptation

ApplicationsDGX agent

arXiv:2604.19488v1 Announce Type: new Abstract: Large language models (LLMs) have achieved substantial advances in logical reasoning, yet they continue to lag behind human-level performance. In-contex

CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation

Model ReleasesDGX agent

arXiv:2604.19636v1 Announce Type: new Abstract: Synthesizing human--object interaction (HOI) videos has broad practical value in e-commerce, digital advertising, and virtual marketing. However, curren

Collaborative Contextual Bayesian Optimization

ApplicationsDGX agent

arXiv:2604.18912v1 Announce Type: new Abstract: Discovering optimal designs through sequential data collection is essential in many real-world applications. While Bayesian Optimization (BO) has achiev

Colour Extraction Pipeline for Odonates using Computer Vision

ResearchDGX agent

arXiv:2604.18725v1 Announce Type: new Abstract: The correlation between insect morphological traits and climate has been documented in physiological studies, but such studies remain limited by the tim

COMODO: Cross-Modal Video-to-IMU Distillation for Efficient Egocentric Human Activity Recognition

Local AiDGX agent

arXiv:2503.07259v2 Announce Type: replace-cross Abstract: The goal of creating intelligent, human-centered wearable systems for continuous activity understanding faces a fundamental trade-off: Egocent

Comparing energy consumption and accuracy in text classification inference

ResearchDGX agent

arXiv:2508.14170v2 Announce Type: replace Abstract: The increasing deployment of large language models (LLMs) in natural language processing (NLP) tasks raises concerns about energy efficiency and sus

Comparison of sEMG Encoding Accuracy Across Speech Modes Using Articulatory and Phoneme Features

ResearchDGX agent

arXiv:2604.18920v1 Announce Type: cross Abstract: We test whether Speech Articulatory Coding (SPARC) features can linearly predict surface electromyography (sEMG) envelopes across aloud, mimed, and su

Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs

Model ReleasesDGX agent

arXiv:2604.18587v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated significant potential in formal theorem proving, yet state-of-the-art performance often necessitates pr

Computational Narrative Understanding for Expressive Text-to-Speech

ResearchDGX agent

arXiv:2509.04072v2 Announce Type: replace-cross Abstract: Recent advances in text-to-speech (TTS) have been driven by large, multi-domain speech corpora, yet the expressive potential of audiobook data

Concept Inconsistency in Dermoscopic Concept Bottleneck Models: A Rough-Set Analysis of the Derm7pt Dataset

Model ReleasesDGX agent

arXiv:2604.19323v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) route predictions exclusively through a clinically grounded concept layer, binding interpretability to concept-label

Conditional Diffusion Modeling with Attention for Probabilistic Battery Capacity Prediction under Real-World Condition

ApplicationsDGX agent

arXiv:2510.17414v2 Announce Type: replace Abstract: Accurate prediction of lithium-ion battery capacity and its associated uncertainty is essential for reliable battery management but remains challeng

Conjuring Semantic Similarity

ResearchDGX agent

arXiv:2410.16431v4 Announce Type: replace Abstract: The semantic similarity between sample expressions measures the distance between their latent 'meaning'. These meanings are themselves typically rep

Construction of Knowledge Graph based on Language Model

ResearchDGX agent

arXiv:2604.19137v1 Announce Type: new Abstract: Knowledge Graph (KG) can effectively integrate valuable information from massive data, and thus has been rapidly developed and widely used in many field

ContextLeak: Auditing Leakage in Private In-Context Learning Methods

TutorialsDGX agent

arXiv:2512.16059v2 Announce Type: replace-cross Abstract: In-Context Learning (ICL) has become a standard technique for adapting Large Language Models (LLMs) to specialized tasks by supplying task-spe

ConvVitMamba: Efficient Multiscale Convolution, Transformer, and Mamba-Based Sequence modelling for Hyperspectral Image Classification

Model ReleasesDGX agent

arXiv:2604.18856v1 Announce Type: new Abstract: Hyperspectral image (HSI) classification remains challenging due to high spectral dimensionality, redundancy, and limited labeled data. Although convolu

Council Mode: Mitigating Hallucination and Bias in LLMs via Multi-Agent Consensus

Model ReleasesDGX agent

arXiv:2604.02923v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly those employing Mixture-of-Experts (MoE) architectures, have achieved remarkable capabilities acros

CounterRefine: Answer-Conditioned Counterevidence Retrieval for Inference-Time Knowledge Repair in Factual Question Answering

Model ReleasesDGX agent

arXiv:2603.16091v2 Announce Type: replace-cross Abstract: In factual question answering, many errors are not failures of access but failures of commitment: the system retrieves relevant evidence, yet

Counting Worlds Branching Time Semantics for post-hoc Bias Mitigation in generative AI

SafetyDGX agent

arXiv:2604.19431v1 Announce Type: cross Abstract: Generative AI systems are known to amplify biases present in their training data. While several inference-time mitigation strategies have been propose

CreatiParser: Generative Image Parsing of Raster Graphic Designs into Editable Layers

SafetyDGX agent

arXiv:2604.19632v1 Announce Type: new Abstract: Graphic design images consist of multiple editable layers, such as text, background, and decorative elements, while most generative models produce raste

Cross-lingual Matryoshka Representation Learning across Speech and Text

ResearchDGX agent

arXiv:2602.19991v2 Announce Type: replace Abstract: Speakers of under-represented languages face both a language barrier, as most online knowledge is in a few dominant languages, and a modality barrie

Cross-Model Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Across Three Large Language Models

Model ReleasesDGX agent

arXiv:2604.19598v1 Announce Type: cross Abstract: This study compared repeated generation consistency of exercise prescription outputs across three large language models (LLMs), specifically GPT-4.1,

CrossPan: A Comprehensive Benchmark for Cross-Sequence Pancreas MRI Segmentation and Generalization

Model ReleasesDGX agent

arXiv:2604.18797v1 Announce Type: new Abstract: Automatic pancreas segmentation is fundamental to abdominal MRI analysis, yet deep learning models trained on one MRI sequence often fail catastrophical

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks

Model ReleasesDGX agent

arXiv:2604.19262v1 Announce Type: cross Abstract: Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities.

Curiosity-Critic: Cumulative Prediction Error Improvement as a Tractable Intrinsic Reward for World Model Training

Local AiDGX agent

arXiv:2604.18701v1 Announce Type: cross Abstract: Local prediction-error-based curiosity rewards focus on the current transition without considering the world model's cumulative prediction error acros

Curvature-Aware PCA with Geodesic Tangent Space Aggregation for Semi-Supervised Learning

SafetyDGX agent

arXiv:2604.18816v1 Announce Type: cross Abstract: Principal Component Analysis (PCA) is a fundamental tool for representation learning, but its global linear formulation fails to capture the structure

← Previous
1…870871872873874…998
Next →