AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Revisiting Non-Verbatim Memorization in Large Language Models: The Role of Entity Surface Forms

DGX agent

arXiv:2604.21882v1 Announce Type: new Abstract: Understanding what kinds of factual knowledge large language models (LLMs) memorize is essential for evaluating their reliability and limitations. Entit

researcharxiv-cs-cl
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

S1-VL: Scientific Multimodal Reasoning Model with Thinking-with-Images

DGX agent

arXiv:2604.21409v1 Announce Type: new Abstract: We present S1-VL, a multimodal reasoning model for scientific domains that natively supports two complementary reasoning paradigms: Scientific Reasoning

tutorialsarxiv-cs-cv
24 Apr 2026
Model Releases

Synthetic Data in Education: Empirical Insights from Traditional Resampling and Deep Generative Models

DGX agent

arXiv:2604.21031v1 Announce Type: cross Abstract: Synthetic data generation offers promise for addressing data scarcity and privacy concerns in educational technology, yet practitioners lack empirical

model-releasesarxiv-cs-ai
24 Apr 2026
Research

UKP_Psycontrol at SemEval-2026 Task 2: Modeling Valence and Arousal Dynamics from Text

DGX agent

arXiv:2604.21534v1 Announce Type: new Abstract: This paper presents our system developed for SemEval-2026 Task 2. The task requires modeling both current affect and short-term affective change in chro

researcharxiv-cs-cl
24 Apr 2026
Safety

A Survey of Scaling in Large Language Model Reasoning

DGX agent

arXiv:2504.02181v2 Announce Type: replace Abstract: The rapid advancements in large Language models (LLMs) have significantly enhanced their reasoning capabilities, driven by various strategies such a

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Cognitive Kernel-Pro: A Framework for Deep Research Agents and Agent Foundation Models Training

DGX agent

arXiv:2508.00414v3 Announce Type: replace Abstract: General AI Agents are increasingly recognized as foundational frameworks for the next generation of artificial intelligence, enabling complex reason

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Improving End-to-End Training of Retrieval-Augmented Generation Models via Joint Stochastic Approximation

DGX agent

arXiv:2508.18168v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) has become a widely recognized paradigm to combine parametric memory with non-parametric memories. An RAG model

researcharxiv-cs-cl
23 Apr 2026
Model Releases

Large Language Models Meet Biomedical Knowledge Graphs for Mechanistically Grounded Therapeutic Prioritization

DGX agent

arXiv:2604.19815v1 Announce Type: new Abstract: Drug repurposing is often framed as a candidate identification task, but existing approaches provide limited guidance for distinguishing biologically pl

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Large language models perceive cities through a culturally uneven baseline

DGX agent

arXiv:2604.20048v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to describe, evaluate and interpret places, yet it remains unclear whether they do so from a cultural

safetyarxiv-cs-cl
23 Apr 2026
Safety

ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation

DGX agent

arXiv:2604.20666v1 Announce Type: cross Abstract: Effective retrieval-augmented generation across bilingual Greek--English applications requires embedding models capable of capturing both domain-speci

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

RareSpot+: A Benchmark, Model, and Active Learning Framework for Small and Rare Wildlife in Aerial Imagery

DGX agent

arXiv:2604.20000v1 Announce Type: new Abstract: Automated wildlife monitoring from aerial imagery is vital for conservation but remains limited by two persistent challenges: the difficulty of detectin

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

Resolving space-sharing conflicts in road user interactions through uncertainty reduction: An active inference-based computational model

DGX agent

arXiv:2604.19838v1 Announce Type: new Abstract: Understanding how road users resolve space-sharing conflicts is important both for traffic safety and the safe deployment of autonomous vehicles. While

safetyarxiv-cs-ai
23 Apr 2026
Research

Saying More Than They Know: A Framework for Quantifying Epistemic-Rhetorical Miscalibration in Large Language Models

DGX agent

arXiv:2604.19768v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic miscalibration with rhetorical intensity not proportionate to epistemic grounding. This study tests th

researcharxiv-cs-ai
23 Apr 2026
Safety

Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models

DGX agent

arXiv:2604.20472v1 Announce Type: cross Abstract: Recent advances in vision-language-action (VLA) models for robotics have highlighted the importance of reliable uncertainty quantification in sequenti

safetyarxiv-cs-lg
23 Apr 2026
Research

Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling

DGX agent

arXiv:2604.01577v2 Announce Type: replace-cross Abstract: We extend the recent latent recurrent modeling to sequential input streams. By interleaving fast, recurrent latent updates with self-organizat

researcharxiv-cs-ai
23 Apr 2026
Research

What Makes a Bacterial Model a Good Reservoir Computer? Predicting Performance from Separability and Similarity

DGX agent

arXiv:2604.19850v1 Announce Type: cross Abstract: Biological systems are promising substrates for computation because they naturally process environmental information through complex internal dynamics

researcharxiv-cs-lg
23 Apr 2026
Research

Benchmarking Vision Foundation Models for Domain-Generalizable Face Anti-Spoofing

DGX agent

arXiv:2604.19196v1 Announce Type: new Abstract: Face Anti-Spoofing (FAS) remains challenging due to the requirement for robust domain generalization across unseen environments. While recent trends lev

researcharxiv-cs-cv
22 Apr 2026
Model Releases

Byzantine-tolerant distributed learning of finite mixture models

DGX agent

arXiv:2407.13980v3 Announce Type: replace-cross Abstract: Traditional statistical methods need to be updated to work with modern distributed data storage paradigms. A common approach is the split-and-

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

Diagnosable ColBERT: Debugging Late-Interaction Retrieval Models Using a Learned Latent Space as Reference

DGX agent

arXiv:2604.19566v1 Announce Type: cross Abstract: Reliable biomedical and clinical retrieval requires more than strong ranking performance: it requires a practical way to find systematic model failure

safetyarxiv-cs-cl
22 Apr 2026
Safety

Fairness Audits of Institutional Risk Models in Deployed ML Pipelines

DGX agent

arXiv:2604.19468v1 Announce Type: cross Abstract: Fairness audits of institutional risk models are critical for understanding how deployed machine learning pipelines allocate resources. Drawing on mul

safetyarxiv-cs-ai
22 Apr 2026
Safety

Persuasion with Large Language Models: A Survey of Empirical Evidence, Study Methodologies, and Ethical Implications

DGX agent

arXiv:2411.06837v2 Announce Type: replace Abstract: The rapid rise of Large Language Models (LLMs) has created new disruptive possibilities for persuasive communication, enabling fully-automated, pers

safetyarxiv-cs-cl
22 Apr 2026
Applications

Reduced-Order Surrogates for Forced Flexible Mesh Coastal-Ocean Models

DGX agent

arXiv:2602.05416v2 Announce Type: replace-cross Abstract: While proper orthogonal decomposition (POD)-based surrogates are widely explored for hydrodynamic applications, the use of Koopman autoencoder

applicationsarxiv-cs-ai
22 Apr 2026
Hardware

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

DGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

hardwarearxiv-cs-ai
22 Apr 2026
Local Ai

StrikeWatch: Wrist-worn Gait Recognition with Compact Time-series Models on Low-power FPGAs

DGX agent

arXiv:2510.24738v2 Announce Type: replace-cross Abstract: Running offers substantial health benefits, but improper gait patterns can lead to injuries, particularly without expert feedback. While prior

local-aiarxiv-cs-lg
22 Apr 2026
Research

A Two-Phase Deep Learning Framework for Adaptive Time-Stepping in High-Speed Flow Modeling

DGX agent

arXiv:2506.07969v2 Announce Type: replace Abstract: We consider the problem of modeling high-speed flows using machine learning methods. While most prior studies focus on low-speed fluid flows in whic

researcharxiv-cs-lg
21 Apr 2026
Applications

Active World-Model with 4D-informed Retrieval for Exploration and Awareness

DGX agent

arXiv:2604.16733v1 Announce Type: new Abstract: Physical awareness, especially in a large and dynamic environment, is shaped by sensing decisions that determine observability across space, time, and s

applicationsarxiv-cs-cv
21 Apr 2026
Model Releases

Bridging the Reasoning Gap in Vietnamese with Small Language Models via Test-Time Scaling

DGX agent

arXiv:2604.17794v1 Announce Type: new Abstract: The democratization of ubiquitous AI hinges on deploying sophisticated reasoning capabilities on resource-constrained devices. However, Small Language M

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Comparing Human and Large Language Model Interpretation of Implicit Information

DGX agent

arXiv:2604.17085v1 Announce Type: new Abstract: The interpretation of implicit meanings is an integral aspect of human communication. However, this framework may not transfer to interactions with Larg

researcharxiv-cs-cl
21 Apr 2026
Research

Comparison Drives Preference: Reference-Aware Modeling for AI-Generated Video Quality Assessment

DGX agent

arXiv:2604.17074v1 Announce Type: new Abstract: The rapid advancement of generative models has led to a growing volume of AI-generated videos, making the automatic quality assessment of such videos in

researcharxiv-cs-cv
21 Apr 2026
Safety

Cross-Modal Attention Analysis and Optimization in Vision-Language Models: A Study on Visual Reliability

DGX agent

arXiv:2604.17217v1 Announce Type: new Abstract: Vision-Language Models (VLMs) achieve strong cross-modal performance, yet recent evidence suggests they over-rely on textual descriptions while under-ut

safetyarxiv-cs-cv
21 Apr 2026
Research

DART: Learning-Enhanced Model Predictive Control for Dual-Arm Non-Prehensile Manipulation

DGX agent

arXiv:2604.17833v1 Announce Type: new Abstract: What appears effortless to a human waiter remains a major challenge for robots. Manipulating objects nonprehensilely on a tray is inherently difficult,

researcharxiv-cs-ro
21 Apr 2026
Research

Data Mixing for Large Language Models Pretraining: A Survey and Outlook

DGX agent

arXiv:2604.16380v1 Announce Type: new Abstract: Large language models (LLMs) rely on pretraining on massive and heterogeneous corpora, where training data composition has a decisive impact on training

researcharxiv-cs-cl
21 Apr 2026
Applications

DexWorldModel: Causal Latent World Modeling towards Automated Learning of Embodied Tasks

DGX agent

arXiv:2604.16484v1 Announce Type: new Abstract: Deploying generative World-Action Models for manipulation is severely bottlenecked by redundant pixel-level reconstruction, O(T) memory scaling, and seq

applicationsarxiv-cs-cv
21 Apr 2026
Safety

Dual Alignment Between Language Model Layers and Human Sentence Processing

DGX agent

arXiv:2604.18563v1 Announce Type: new Abstract: A recent study (Kuribayashi et al., 2025) has shown that human sentence processing behavior, typically measured on syntactically unchallenging construct

safetyarxiv-cs-cl
21 Apr 2026
Research

Dynamic Eraser for Guided Concept Erasure in Diffusion Models

DGX agent

arXiv:2604.16483v1 Announce Type: new Abstract: Concept erasure in Text-To-Image (T2I) diffusion models is vital for safe content generation, but existing inference-time methods face significant limit

researcharxiv-cs-cv
21 Apr 2026
Safety

Efficient Diffusion Models under Nonconvex Equality and Inequality constraints via Landing

DGX agent

arXiv:2604.17838v1 Announce Type: new Abstract: Generative modeling within constrained sets is essential for scientific and engineering applications involving physical, geometric, or safety requiremen

safetyarxiv-cs-lg
21 Apr 2026
Research

Efficient Low-Resource Language Adaptation via Multi-Source Dynamic Logit Fusion

DGX agent

arXiv:2604.18106v1 Announce Type: new Abstract: Adapting large language models (LLMs) to low-resource languages (LRLs) is constrained by the scarcity of task data and computational resources. Although

researcharxiv-cs-cl
21 Apr 2026
Safety

Fairness Constraints in High-Dimensional Generalized Linear Models

DGX agent

arXiv:2604.16610v1 Announce Type: cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness

safetyarxiv-cs-lg
21 Apr 2026
Local Ai

How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them

DGX agent

arXiv:2604.17105v1 Announce Type: new Abstract: Tokenization is the first step in every language model (LM), yet it never takes the sounds of words into account. We investigate how tokenization influe

local-aiarxiv-cs-cl
21 Apr 2026
Agents

Learning to Trade Like an Expert: Cognitive Fine-Tuning for Stable Financial Reasoning in Language Models

DGX agent

arXiv:2604.16862v1 Announce Type: new Abstract: Recent deployments of large language models (LLMs) as autonomous trading agents raise questions about whether financial decision-making competence gener

agentsarxiv-cs-lg
21 Apr 2026
Model Releases

Leveraging Large Language Models for Sarcastic Speech Annotation in Sarcasm Detection

DGX agent

arXiv:2506.00955v2 Announce Type: replace Abstract: Sarcasm fundamentally alters meaning through tone and context, yet detecting it in speech remains a challenge due to data scarcity. In addition, exi

model-releasesarxiv-cs-cl
21 Apr 2026
Research

LLM-AUG: Robust Wireless Data Augmentation with In-Context Learning in Large Language Models

DGX agent

arXiv:2604.17770v1 Announce Type: new Abstract: Data scarcity remains a fundamental bottleneck in applying deep learning to wireless communication problems, particularly in scenarios where collecting

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Low-rank Orthogonalization for Large-scale Matrix Optimization with Applications to Foundation Model Training

DGX agent

arXiv:2509.11983v2 Announce Type: replace Abstract: Neural network (NN) training is inherently a large-scale matrix optimization problem, yet the matrix structure of NN parameters has long been overlo

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

DGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

DGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

safetyarxiv-cs-cl
21 Apr 2026
Applications

Modeling Human Perspectives with Socio-Demographic Representations

DGX agent

arXiv:2604.18069v1 Announce Type: new Abstract: Humans often hold different perspectives on the same issues. In many NLP tasks, annotation disagreement can reflect valid subjective perspectives. Model

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

Modeling Multiple Support Strategies within a Single Turn for Emotional Support Conversations

DGX agent

arXiv:2604.17972v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) aims to assist individuals experiencing distress by generating empathetic and supportive dialogue. While prior work

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

NaviFormer: A Deep Reinforcement Learning Transformer-like Model to Holistically Solve the Navigation Problem

DGX agent

arXiv:2604.16967v1 Announce Type: new Abstract: Path planning is usually solved by addressing either the (high-level) route planning problem (waypoint sequencing to achieve the final goal) or the (low

applicationsarxiv-cs-ro
21 Apr 2026
← Previous
1…141142143144145…1030
Next →