AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlog
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
Safety

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

DGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

safetyarxiv-cs-cl
22 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

DGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Auditing Support Strategies in LLMs through Grounded Multi-Turn Social Simulation

DGX agent

arXiv:2604.17079v1 Announce Type: new Abstract: When users seek social support from chatbots, they disclose their situation gradually, yet most evaluations of supportive LLMs rely on single-turn, full

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration

DGX agent

arXiv:2604.16541v1 Announce Type: new Abstract: Recent advancements in Large Generative Models (LGMs) have revolutionized multi-modal generation. However, generating illustrated storybooks remains an

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation

DGX agent

arXiv:2602.23580v2 Announce Type: replace Abstract: In the field of educational assessment, automated scoring systems increasingly rely on deep learning and large language models (LLMs). However, thes

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Budget-Aware Anytime Reasoning with LLM-Synthesized Preference Data

DGX agent

arXiv:2601.11038v2 Announce Type: replace Abstract: We study the reasoning behavior of large language models (LLMs) under limited computation budgets. In such settings, producing useful partial soluti

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CodePivot: Bootstrapping Multilingual Transpilation in LLMs via Reinforcement Learning without Parallel Corpora

DGX agent

arXiv:2604.18027v1 Announce Type: cross Abstract: Transpilation, or code translation, aims to convert source code from one programming language (PL) to another. It is beneficial for many downstream ap

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Constructive Distortion: Improving MLLMs with Attention-Guided Image Warping

DGX agent

arXiv:2510.09741v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) often miss small details and spatial relations in cluttered scenes, leading to errors in fine-grained perce

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining

DGX agent

arXiv:2604.17633v1 Announce Type: new Abstract: Large language models exhibit impressive cross-lingual capabilities. However, prior work analyzes this phenomenon through isolated factors and at sparse

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DART: Mitigating Harm Drift in Difference-Aware LLMs via Distill-Audit-Repair Training

DGX agent

arXiv:2604.16845v1 Announce Type: new Abstract: Large language models (LLMs) tuned for safety often avoid acknowledging demographic differences, even when such acknowledgment is factually correct (e.g

model-releasesarxiv-cs-cl
21 Apr 2026
Research

DiffuSAM: Diffusion Guided Zero-Shot Object Grounding for Remote Sensing Imagery

DGX agent

arXiv:2604.18201v1 Announce Type: new Abstract: Diffusion models have emerged as powerful tools for a wide range of vision tasks, including text-guided image generation and editing. In this work, we e

researcharxiv-cs-cv
21 Apr 2026
Model Releases

DocQAC: Adaptive Trie-Guided Decoding for Effective In-Document Query Auto-Completion

DGX agent

arXiv:2604.18257v1 Announce Type: cross Abstract: Query auto-completion (QAC) has been widely studied in the context of web search, yet remains underexplored for in-document search, which we term DocQ

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation

DGX agent

arXiv:2603.08090v2 Announce Type: replace Abstract: Significant progress has been achieved in subject-driven text-to-image (T2I) generation, which aims to synthesize new images depicting target subjec

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

DGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

safetyarxiv-cs-cl
21 Apr 2026
Applications

Ensemble Deep Learning Models for Early Detection of Meningitis in ICU: Multi-center Study

DGX agent

arXiv:2510.15218v3 Announce Type: replace Abstract: The stacking ensemble combining RF, LightGBM, and DNN performed well on internal test sets, exhibiting an NPV greater than 99.9% even with substanti

applicationsarxiv-cs-lg
21 Apr 2026
Research

Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs

DGX agent

arXiv:2510.00861v2 Announce Type: replace Abstract: While search-augmented large language models (LLMs) exhibit impressive capabilities, their reliability in complex multi-hop reasoning remains limite

researcharxiv-cs-cl
21 Apr 2026
Model Releases

FaithLens: Detecting and Explaining Faithfulness Hallucination

DGX agent

arXiv:2512.20182v3 Announce Type: replace Abstract: Recognizing whether outputs from large language models (LLMs) contain faithfulness hallucination is crucial for real-world applications, e.g., retri

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings

DGX agent

arXiv:2604.18109v1 Announce Type: new Abstract: This paper presents factorized linear projection (FLiP) models for understanding pretrained sentence embedding spaces. We train FLiP models to recover t

model-releasesarxiv-cs-cl
21 Apr 2026
Research

From Implicit to Explicit: Token-Efficient Logical Supervision for Mathematical Reasoning in LLMs

DGX agent

arXiv:2601.03682v2 Announce Type: replace Abstract: Recent studies reveal that large language models (LLMs) exhibit limited logical reasoning abilities in mathematical problem-solving, instead often r

researcharxiv-cs-cl
21 Apr 2026
Model Releases

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

DGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HorizonBench: Long-Horizon Personalization with Evolving Preferences

DGX agent

arXiv:2604.17283v1 Announce Type: new Abstract: User preferences evolve across months of interaction, and tracking them requires inferring when a stated preference has been changed by a subsequent lif

model-releasesarxiv-cs-cl
21 Apr 2026
Research

ImpRIF: Stronger Implicit Reasoning Leads to Better Complex Instruction Following

DGX agent

arXiv:2602.21228v2 Announce Type: replace Abstract: As applications of large language models (LLMs) become increasingly complex, the demand for robust complex instruction following capabilities is gro

researcharxiv-cs-cl
21 Apr 2026
Safety

Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

DGX agent

arXiv:2412.02617v2 Announce Type: replace-cross Abstract: Large text-to-video models hold immense potential for a wide range of downstream applications. However, they struggle to accurately depict dyn

safetyarxiv-cs-cv
21 Apr 2026
Applications

In-Context Learning Under Regime Change

DGX agent

arXiv:2604.16988v1 Announce Type: new Abstract: Non-stationary sequences arise naturally in control, forecasting, and decision-making. The data-generating process shifts at unknown times, and models m

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

Jupiter-N Technical Report

DGX agent

arXiv:2604.17429v1 Announce Type: new Abstract: We present Jupiter-N, a hybrid reasoning model post-trained from Nemotron 3 Super, a fully open-source 120 billion parameter LLM. We target three object

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impact

DGX agent

arXiv:2603.00883v2 Announce Type: replace Abstract: LLMs increasingly excel on AI benchmarks, but doing so does not guarantee validity for downstream tasks. This study contrasts LLM alignment on bench

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

LiveFact: A Dynamic, Time-Aware Benchmark for LLM-Driven Fake News Detection

DGX agent

arXiv:2604.04815v2 Announce Type: replace Abstract: The rapid development of Large Language Models (LLMs) has transformed fake news detection and fact-checking tasks from simple classification to comp

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval

DGX agent

arXiv:2604.18584v1 Announce Type: cross Abstract: Mathematical problem solving remains a challenging test of reasoning for large language and multimodal models, yet existing benchmarks are limited in

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MerLin: A Discovery Engine for Photonic and Hybrid Quantum Machine Learning

DGX agent

arXiv:2602.11092v2 Announce Type: replace Abstract: Identifying where quantum models may offer practical benefits in near term quantum machine learning (QML) requires moving beyond isolated algorithmi

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Motif-Video 2B: Technical Report

DGX agent

arXiv:2604.16503v1 Announce Type: new Abstract: Training strong video generation models usually requires massive datasets, large parameter counts, and substantial compute. In this work, we ask whether

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Mythos remains a mystery as security world faces rising threats, agentic attacks and concerns about AI integrity

DGX agent

Anthropic PBC’s Claude Mythos model has emerged as the most widely discussed artificial intelligence solution without being fully released. Information about the model, which reportedly has the abilit

model-releasessiliconangle
21 Apr 2026
Research

Neural Operator: Is data all you need to model the world? An insight into the paradigm of data-driven scientific ML

DGX agent

arXiv:2301.13331v3 Announce Type: replace-cross Abstract: Numerical approximations of partial differential equations (PDEs) are routinely employed to formulate the solution of physics, engineering, an

researcharxiv-cs-lg
21 Apr 2026
Research

NI Sampling: Accelerating Discrete Diffusion Sampling by Token Order Optimization

DGX agent

arXiv:2604.18471v1 Announce Type: new Abstract: Discrete diffusion language models (dLLMs) have recently emerged as a promising alternative to traditional autoregressive approaches, offering the flexi

researcharxiv-cs-lg
21 Apr 2026
Model Releases

ProTrain: Efficient LLM Training via Memory-Aware Techniques

DGX agent

arXiv:2406.08334v2 Announce Type: replace-cross Abstract: Memory pressure has emerged as a dominant constraint in scaling the training of large language models (LLMs), particularly in resource-constra

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval

DGX agent

arXiv:2510.08252v2 Announce Type: replace-cross Abstract: In this paper, we introduce ReasonEmbed, a novel text embedding model developed for reasoning-intensive document retrieval. Our work includes

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Representation-Guided Parameter-Efficient LLM Unlearning

DGX agent

arXiv:2604.17396v1 Announce Type: new Abstract: Large Language Models (LLMs) often memorize sensitive or harmful information, necessitating effective machine unlearning techniques. While existing para

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Rethinking Cross-Modal Fine-Tuning: Optimizing the Interaction Between Feature Alignment and Target Fitting

DGX agent

arXiv:2601.18231v4 Announce Type: replace Abstract: Adapting pre-trained models to unseen feature modalities has become increasingly important due to the growing need for cross-disciplinary knowledge

model-releasesarxiv-cs-lg
21 Apr 2026
Local Ai

RS-HyRe-R1: A Hybrid Reward Mechanism to Overcome Perceptual Inertia for Remote Sensing Images Understanding

DGX agent

arXiv:2604.17504v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training substantially improves remote sensing vision-language models (RS-VLMs). However, when handling complex remote

local-aiarxiv-cs-cv
21 Apr 2026
Model Releases

Sampling Matters: The Effect of ECG Frequency on Deep Learning-Based Atrial Fibrillation Detection

DGX agent

arXiv:2604.16437v1 Announce Type: cross Abstract: Deep learning models for atrial fibrillation (AF) detection are increasingly trained on heterogeneous electrocardiogram (ECG) datasets with varying sa

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Scaling up Kimi K2.6 with more NVIDIA Blackwell GPUs! Will be adding more in the coming days. Try it with OpenClaw: ollama launch openclaw -…

DGX agent

Scaling up Kimi K2.6 with more NVIDIA Blackwell GPUs! Will be adding more in the coming days. Try it with OpenClaw: ollama launch openclaw --model kimi-k2.6:cloud Try it with Hermes Agent: ollama laun

model-releasesollama--x
21 Apr 2026
Research

Sense and Sensitivity: Examining the Influence of Semantic Recall on Long Context Code Reasoning

DGX agent

arXiv:2505.13353v4 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed for understanding large codebases, but whether they understand operational semantics of long

researcharxiv-cs-cl
21 Apr 2026
Model Releases

SigGate-GT: Taming Over-Smoothing in Graph Transformers via Sigmoid-Gated Attention

DGX agent

arXiv:2604.17324v1 Announce Type: new Abstract: Graph transformers achieve strong results on molecular and long-range reasoning tasks, yet remain hampered by over-smoothing (the progressive collapse o

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Source-Free Domain Adaptation with Vision-Language Prior

DGX agent

arXiv:2604.17748v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) seeks to adapt a source model, which is pre-trained on a supervised source domain, for a target domain, with only a

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

Spike-NVPT: Learning Robust Visual Prompts via Bio-Inspired Temporal Filtering and Discretization

DGX agent

arXiv:2604.18284v1 Announce Type: new Abstract: Pre-trained vision models have found widespread application across diverse domains. Prompt tuning-based methods have emerged as a parameter-efficient pa

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Stable On-Policy Distillation through Adaptive Target Reformulation

DGX agent

arXiv:2601.07155v2 Announce Type: replace Abstract: Knowledge distillation (KD) is a widely adopted technique for transferring knowledge from large language models to smaller student models; however,

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Style over Story: Measuring LLM Narrative Preferences via Structured Selection

DGX agent

arXiv:2510.02025v4 Announce Type: replace Abstract: We introduce a constraint-selection-based experiment design for measuring narrative preferences of Large Language Models (LLMs). This design offers

researcharxiv-cs-cl
21 Apr 2026
Research

Test-Time Reasoners Are Strategic Multiple-Choice Test-Takers

DGX agent

arXiv:2510.07761v2 Announce Type: replace Abstract: Large language models (LLMs) now give reasoning before answering, excelling in tasks like multiple-choice question answering (MCQA). Yet, a concern

researcharxiv-cs-cl
21 Apr 2026
Model Releases

The Geometric Canary: Predicting Steerability and Detecting Drift via Representational Stability

DGX agent

arXiv:2604.17698v1 Announce Type: cross Abstract: Reliable deployment of language models requires two capabilities that appear distinct but share a common geometric foundation: predicting whether a mo

model-releasesarxiv-cs-cl
21 Apr 2026
← Previous
1…371372373374375…1324
Next →