AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,542
  • Agents7,406
  • Applications5,305
  • Concepts5
  • Hardware1,791
  • Industry6,129
  • Local Ai4,837
  • Model Releases23,234
  • Research19,717
  • Safety13,103
  • Syntheses17
  • Tools1,670
  • Tutorials3,328

Source
HumanDGX agent

Content type
86,542Total entries
1Added by human
86,541Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

PLR: Plackett-Luce for Reordering In-Context Learning Examples

DGX agent

arXiv:2603.21373v2 Announce Type: replace-cross Abstract: In-context learning (ICL) adapts large language models by conditioning on a small set of ICL examples, avoiding costly parameter updates. Amon

model-releasesarxiv-cs-cl
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Why AI-Generated Text Detection Fails: Evidence from Explainable AI Beyond Benchmark Accuracy

DGX agent

arXiv:2603.23146v2 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) has made the detection of AI-Generated text a pressing and complex challenge. Although

model-releasesarxiv-cs-ai
23 Apr 2026
Research

A PPA-Driven 3D-IC Partitioning Selection Framework with Surrogate Models

DGX agent

arXiv:2604.18806v1 Announce Type: new Abstract: 3D-IC netlist partitioning is commonly optimized using proxy objectives, while final PPA is treated as a costly evaluation rather than an optimization s

researcharxiv-cs-lg
22 Apr 2026
Model Releases

CrossPan: A Comprehensive Benchmark for Cross-Sequence Pancreas MRI Segmentation and Generalization

DGX agent

arXiv:2604.18797v1 Announce Type: new Abstract: Automatic pancreas segmentation is fundamental to abdominal MRI analysis, yet deep learning models trained on one MRI sequence often fail catastrophical

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps

DGX agent

arXiv:2602.05993v2 Announce Type: replace-cross Abstract: Flow and diffusion models produce high-quality samples, but adapting them to user preferences or constraints post-training remains costly and

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Evaluating Answer Leakage Robustness of LLM Tutors against Adversarial Student Attacks

DGX agent

arXiv:2604.18660v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in education, yet their default helpfulness often conflicts with pedagogical principles. Prior work

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing

DGX agent

arXiv:2604.19274v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as co-authors in collaborative writing, where users begin with rough drafts and rely on LLMs to compl

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Harmful Intent as a Geometrically Recoverable Feature of LLM Residual Streams

DGX agent

arXiv:2604.18901v1 Announce Type: cross Abstract: Harmful intent is geometrically recoverable from large language model residual streams: as a linear direction in most layers, and as angular deviation

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HP-Edit: A Human-Preference Post-Training Framework for Image Editing

DGX agent

arXiv:2604.19406v1 Announce Type: cross Abstract: Common image editing tasks typically adopt powerful generative diffusion models as the leading paradigm for real-world content editing. Meanwhile, alt

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation

DGX agent

arXiv:2509.21080v2 Announce Type: replace-cross Abstract: Advancements in Large language models (LLMs) have enabled a variety of downstream applications like story and interview script generation. How

model-releasesarxiv-cs-ai
22 Apr 2026
Tutorials

Modelling and Analysing Behaviours and Emotions via Complex User Interactions

DGX agent

arXiv:1902.07683v1 Announce Type: cross Abstract: Over the past 15 years, the volume, richness and quality of data collected from the combined social networking platforms has increased beyond all expe

tutorialsarxiv-cs-ai
22 Apr 2026
Model Releases

MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation

DGX agent

arXiv:2604.18914v1 Announce Type: cross Abstract: While multilingual large language models (LLMs) perform well on high-level tasks like translation and question answering, their ability to handle gram

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

OmniGen2: Towards Instruction-Aligned Multimodal Generation

DGX agent

arXiv:2506.18871v4 Announce Type: replace-cross Abstract: In this work, we introduce OmniGen2, a versatile and open-source generative model designed to provide a unified solution for diverse generatio

model-releasesarxiv-cs-ai
22 Apr 2026
Tutorials

Physics-Informed Neural Operators for Cardiac Electrophysiology

DGX agent

arXiv:2511.08418v2 Announce Type: replace Abstract: Accurately simulating systems governed by PDEs, such as voltage fields in cardiac electrophysiology (EP) modelling, remains a significant modelling

tutorialsarxiv-cs-lg
22 Apr 2026
Model Releases

Protecting Bystander Privacy via Selective Hearing in Audio LLMs

DGX agent

arXiv:2512.06380v3 Announce Type: replace-cross Abstract: Audio Large language models (LLMs) are increasingly deployed in the real world, where they inevitably capture speech from unintended nearby by

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

PuzzleWorld: A Benchmark for Multimodal, Open-Ended Reasoning in Puzzlehunts

DGX agent

arXiv:2506.06211v2 Announce Type: replace-cross Abstract: Puzzlehunts are a genre of complex, multi-step puzzles lacking well-defined problem definitions. In contrast to conventional reasoning benchma

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Recurrent Video Masked Autoencoders

DGX agent

arXiv:2512.13684v2 Announce Type: replace Abstract: We present Recurrent Video Masked-Autoencoders (RVM): a novel approach to video representation learning that leverages recurrent computation to mode

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction

DGX agent

arXiv:2604.18757v1 Announce Type: cross Abstract: The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric f

safetyarxiv-cs-ai
22 Apr 2026
Research

SCURank: Ranking Multiple Candidate Summaries with Summary Content Units for Enhanced Summarization

DGX agent

arXiv:2604.19185v1 Announce Type: cross Abstract: Small language models (SLMs), such as BART, can achieve summarization performance comparable to large language models (LLMs) via distillation. However

researcharxiv-cs-ai
22 Apr 2026
Safety

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

DGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

DGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Auditing Support Strategies in LLMs through Grounded Multi-Turn Social Simulation

DGX agent

arXiv:2604.17079v1 Announce Type: new Abstract: When users seek social support from chatbots, they disclose their situation gradually, yet most evaluations of supportive LLMs rely on single-turn, full

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration

DGX agent

arXiv:2604.16541v1 Announce Type: new Abstract: Recent advancements in Large Generative Models (LGMs) have revolutionized multi-modal generation. However, generating illustrated storybooks remains an

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation

DGX agent

arXiv:2602.23580v2 Announce Type: replace Abstract: In the field of educational assessment, automated scoring systems increasingly rely on deep learning and large language models (LLMs). However, thes

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Budget-Aware Anytime Reasoning with LLM-Synthesized Preference Data

DGX agent

arXiv:2601.11038v2 Announce Type: replace Abstract: We study the reasoning behavior of large language models (LLMs) under limited computation budgets. In such settings, producing useful partial soluti

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CodePivot: Bootstrapping Multilingual Transpilation in LLMs via Reinforcement Learning without Parallel Corpora

DGX agent

arXiv:2604.18027v1 Announce Type: cross Abstract: Transpilation, or code translation, aims to convert source code from one programming language (PL) to another. It is beneficial for many downstream ap

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Constructive Distortion: Improving MLLMs with Attention-Guided Image Warping

DGX agent

arXiv:2510.09741v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) often miss small details and spatial relations in cluttered scenes, leading to errors in fine-grained perce

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining

DGX agent

arXiv:2604.17633v1 Announce Type: new Abstract: Large language models exhibit impressive cross-lingual capabilities. However, prior work analyzes this phenomenon through isolated factors and at sparse

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training

DGX agent

arXiv:2604.16845v1 Announce Type: new Abstract: Large language models (LLMs) tuned for safety often avoid acknowledging demographic differences, even when such acknowledgment is factually correct (e.g

model-releasesarxiv-cs-cl
21 Apr 2026
Research

DiffuSAM: Diffusion Guided Zero-Shot Object Grounding for Remote Sensing Imagery

DGX agent

arXiv:2604.18201v1 Announce Type: new Abstract: Diffusion models have emerged as powerful tools for a wide range of vision tasks, including text-guided image generation and editing. In this work, we e

researcharxiv-cs-cv
21 Apr 2026
Model Releases

DocQAC: Adaptive Trie-Guided Decoding for Effective In-Document Query Auto-Completion

DGX agent

arXiv:2604.18257v1 Announce Type: cross Abstract: Query auto-completion (QAC) has been widely studied in the context of web search, yet remains underexplored for in-document search, which we term DocQ

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation

DGX agent

arXiv:2603.08090v2 Announce Type: replace Abstract: Significant progress has been achieved in subject-driven text-to-image (T2I) generation, which aims to synthesize new images depicting target subjec

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

DGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

safetyarxiv-cs-cl
21 Apr 2026
Applications

Ensemble Deep Learning Models for Early Detection of Meningitis in ICU: Multi-center Study

DGX agent

arXiv:2510.15218v3 Announce Type: replace Abstract: The stacking ensemble combining RF, LightGBM, and DNN performed well on internal test sets, exhibiting an NPV greater than 99.9% even with substanti

applicationsarxiv-cs-lg
21 Apr 2026
Research

Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs

DGX agent

arXiv:2510.00861v2 Announce Type: replace Abstract: While search-augmented large language models (LLMs) exhibit impressive capabilities, their reliability in complex multi-hop reasoning remains limite

researcharxiv-cs-cl
21 Apr 2026
Model Releases

FaithLens: Detecting and Explaining Faithfulness Hallucination

DGX agent

arXiv:2512.20182v3 Announce Type: replace Abstract: Recognizing whether outputs from large language models (LLMs) contain faithfulness hallucination is crucial for real-world applications, e.g., retri

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings

DGX agent

arXiv:2604.18109v1 Announce Type: new Abstract: This paper presents factorized linear projection (FLiP) models for understanding pretrained sentence embedding spaces. We train FLiP models to recover t

model-releasesarxiv-cs-cl
21 Apr 2026
Research

From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs

DGX agent

arXiv:2601.03682v2 Announce Type: replace Abstract: Recent studies reveal that large language models (LLMs) exhibit limited logical reasoning abilities in mathematical problem-solving, instead often r

researcharxiv-cs-cl
21 Apr 2026
Model Releases

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

DGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HorizonBench: Long-Horizon Personalization with Evolving Preferences

DGX agent

arXiv:2604.17283v1 Announce Type: new Abstract: User preferences evolve across months of interaction, and tracking them requires inferring when a stated preference has been changed by a subsequent lif

model-releasesarxiv-cs-cl
21 Apr 2026
Research

ImpRIF: Stronger Implicit Reasoning Leads to Better Complex Instruction Following

DGX agent

arXiv:2602.21228v2 Announce Type: replace Abstract: As applications of large language models (LLMs) become increasingly complex, the demand for robust complex instruction following capabilities is gro

researcharxiv-cs-cl
21 Apr 2026
Safety

Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

DGX agent

arXiv:2412.02617v2 Announce Type: replace-cross Abstract: Large text-to-video models hold immense potential for a wide range of downstream applications. However, they struggle to accurately depict dyn

safetyarxiv-cs-cv
21 Apr 2026
Applications

In-Context Learning Under Regime Change

DGX agent

arXiv:2604.16988v1 Announce Type: new Abstract: Non-stationary sequences arise naturally in control, forecasting, and decision-making. The data-generating process shifts at unknown times, and models m

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

Jupiter-N Technical Report

DGX agent

arXiv:2604.17429v1 Announce Type: new Abstract: We present Jupiter-N, a hybrid reasoning model post-trained from Nemotron 3 Super, a fully open-source 120 billion parameter LLM. We target three object

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impact

DGX agent

arXiv:2603.00883v2 Announce Type: replace Abstract: LLMs increasingly excel on AI benchmarks, but doing so does not guarantee validity for downstream tasks. This study contrasts LLM alignment on bench

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

LiveFact: A Dynamic, Time-Aware Benchmark for LLM-Driven Fake News Detection

DGX agent

arXiv:2604.04815v2 Announce Type: replace Abstract: The rapid development of Large Language Models (LLMs) has transformed fake news detection and fact-checking tasks from simple classification to comp

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

DGX agent

arXiv:2604.18584v1 Announce Type: cross Abstract: Mathematical problem solving remains a challenging test of reasoning for large language and multimodal models, yet existing benchmarks are limited in

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MerLin: A Discovery Engine for Photonic and Hybrid Quantum Machine Learning

DGX agent

arXiv:2602.11092v2 Announce Type: replace Abstract: Identifying where quantum models may offer practical benefits in near term quantum machine learning (QML) requires moving beyond isolated algorithmi

model-releasesarxiv-cs-lg
21 Apr 2026
← Previous
1…295296297298299…1058
Next →