AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Research

Rethinking Layer Redundancy in Large Language Models: Calibration Objectives and Search for Depth Pruning

DGX agent

arXiv:2604.24938v1 Announce Type: cross Abstract: Depth pruning improves the inference efficiency of large language models by removing Transformer blocks. Prior work has focused on importance criteria

researcharxiv-cs-cl
29 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

DGX agent

arXiv:2406.06587v2 Announce Type: replace Abstract: Aligning large language models (LLMs) behaviour with human intent is critical for future AI. An important yet often overlooked aspect of this alignm

safetyarxiv-cs-cl
29 Apr 2026
Research

Vision SmolMamba: Spike-Guided Token Pruning for Energy-Efficient Spiking State-Space Vision Models

DGX agent

arXiv:2604.25570v1 Announce Type: new Abstract: Spiking Transformers have shown strong potential for long-range visual modeling through spike-driven self-attention. However, their quadratic token inte

researcharxiv-cs-cv
29 Apr 2026
Model Releases

A Theoretical Framework for Auxiliary-Loss-Free Load Balancing of Sparse Mixture-of-Experts in Large-Scale AI Models

DGX agent

arXiv:2512.03915v3 Announce Type: replace-cross Abstract: In large-scale AI training, Sparse Mixture-of-Experts (s-MoE) layers enable scaling by activating only a small subset of experts per token. An

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model

DGX agent

arXiv:2508.06206v4 Announce Type: replace-cross Abstract: Affordance grounding focuses on predicting the specific regions of objects that are associated with the actions to be performed by robots. It

model-releasesarxiv-cs-cv
28 Apr 2026
Local Ai

Breaking the Resource Wall: Geometry-Guided Sequence Modeling for Efficient Semantic Segmentation

DGX agent

arXiv:2604.23399v1 Announce Type: new Abstract: High-performance semantic segmentation has achieved significant progress in recent years, often driven by increasingly large backbones and higher comput

local-aiarxiv-cs-cv
28 Apr 2026
Research

Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation

DGX agent

arXiv:2604.24361v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance in general machine translation, yet their ability in culture-aware scenarios remains poorl

researcharxiv-cs-cl
28 Apr 2026
Tutorials

DeepCausalMMM: A Deep Learning Framework for Marketing Mix Modeling with Causal Structure Learning

DGX agent

arXiv:2510.13087v3 Announce Type: replace Abstract: Marketing Mix Modeling (MMM) estimates the impact of marketing activities on business outcomes such as sales or revenue. Traditional MMM approaches

tutorialsarxiv-cs-lg
28 Apr 2026
Applications

DualGuard: Dual-stream Large Language Model Watermarking Defense against Paraphrase and Spoofing Attack

DGX agent

arXiv:2512.16182v2 Announce Type: replace-cross Abstract: With the rapid development of cloud-based services, large language models have become increasingly accessible through various web platforms. H

applicationsarxiv-cs-cl
28 Apr 2026
Research

Emotion-Conditioned Short-Horizon Human Pose Forecasting with a Lightweight Predictive World Model

DGX agent

arXiv:2604.23532v1 Announce Type: cross Abstract: Short-term human pose prediction plays a crucial role in interactive systems, assistive robots, and emotion-aware human-computer interaction[1-3]. Whi

researcharxiv-cs-ai
28 Apr 2026
Model Releases

FAIR_XAI: Improving Multimodal Foundation Model Fairness via Explainability for Wellbeing Assessment

DGX agent

arXiv:2604.23786v1 Announce Type: new Abstract: In recent years, the integration of multimodal machine learning in wellbeing assessment has offered transformative potential for monitoring mental healt

model-releasesarxiv-cs-ai
28 Apr 2026
Hardware

FreeScale: Distributed Training for Sequence Recommendation Models with Minimal Scaling Cost

DGX agent

arXiv:2604.24073v1 Announce Type: cross Abstract: Modern industrial Deep Learning Recommendation Models typically extract user preferences through the analysis of sequential interaction histories, sub

hardwarearxiv-cs-ai
28 Apr 2026
Tutorials

LunarDepthNet: Generation of Digital Elevation Models using Deep Learning and Monocular Satellite Images

DGX agent

arXiv:2604.22848v1 Announce Type: new Abstract: Recent times have seen an increase in demand of high quality Digital Elevation Models (DEMs) for the lunar surface, because they are highly important fo

tutorialsarxiv-cs-cv
28 Apr 2026
Model Releases

MetaGAI: A Large-Scale and High-Quality Benchmark for Generative AI Model and Data Card Generation

DGX agent

arXiv:2604.23539v1 Announce Type: new Abstract: The rapid proliferation of Generative AI necessitates rigorous documentation standards for transparency and governance. However, manual creation of Mode

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Modeling Induced Pleasure through Cognitive Appraisal Prediction via Multimodal Fusion

DGX agent

arXiv:2604.23753v1 Announce Type: new Abstract: Multimodal affective computing analyzes user-generated social media content to predict emotional states. However, a critical gap remains in understandin

researcharxiv-cs-ai
28 Apr 2026
Applications

Modular Sensory Stream for Integrating Physical Feedback in Vision-Language-Action Models

DGX agent

arXiv:2604.23272v1 Announce Type: new Abstract: Humans understand and interact with the real world by relying on diverse physical feedback beyond visual perception. Motivated by this, recent approache

applicationsarxiv-cs-ro
28 Apr 2026
Model Releases

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

DGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Reflective Flow Sampling Enhancement

DGX agent

arXiv:2603.06165v2 Announce Type: replace-cross Abstract: The growing demand for text-to-image generation has led to rapid advances in generative modeling. Recently, text-to-image diffusion models tra

safetyarxiv-cs-ai
28 Apr 2026
Safety

RL Token: Bootstrapping Online RL with Vision-Language-Action Models

DGX agent

arXiv:2604.23073v1 Announce Type: new Abstract: Vision-language-action (VLA) models can learn to perform diverse manipulation skills 'out of the box,' but achieving the precision and speed that real-w

safetyarxiv-cs-lg
28 Apr 2026
Agents

SCRIBE: Structured Mid-Level Supervision for Tool-Using Language Models

DGX agent

arXiv:2601.03555v2 Announce Type: replace Abstract: Training reliable tool-augmented agents remains a significant challenge, largely due to the difficulty of credit assignment in multi-step reasoning.

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

DGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

model-releasesarxiv-cs-cl
28 Apr 2026
Safety

Seer: Language Instructed Video Prediction with Latent Diffusion Models

DGX agent

arXiv:2303.14897v4 Announce Type: replace Abstract: Imagining the future trajectory is the key for robots to make sound planning and successfully reach their goals. Therefore, text-conditioned video p

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

DGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

DGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

model-releasesarxiv-cs-cl
28 Apr 2026
Agents

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

DGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

agentsarxiv-cs-cl
28 Apr 2026
Model Releases

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

DGX agent

arXiv:2508.09603v2 Announce Type: replace Abstract: Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing

model-releasesarxiv-cs-cl
28 Apr 2026
Applications

Training Machine Learning Models on Encrypted Data: A Privacy-Preserving Framework using Homomorphic Encryption

DGX agent

arXiv:2604.23245v1 Announce Type: cross Abstract: The use of Machine Learning (ML) for data-driven decision-making often relies on access to sensitive datasets, which introduces privacy challenges. Tr

applicationsarxiv-cs-ai
28 Apr 2026
Applications

Variational Grey-Box Dynamics Matching

DGX agent

arXiv:2602.17477v3 Announce Type: replace Abstract: Deep generative models such as flow matching and diffusion models have shown great potential in learning complex distributions and dynamical systems

applicationsarxiv-cs-lg
28 Apr 2026
Research

VS-DDPM: Efficient Low-Cost Diffusion Model for Medical Modality Translation

DGX agent

arXiv:2604.22942v1 Announce Type: cross Abstract: Diffusion models produce high-quality synthetic data but suffer from slow inference. We propose 3D Variable-Step Denoising Diffusion Probabilistic Mod

researcharxiv-cs-ai
28 Apr 2026
Research

Weakly Supervised Multicenter Nancy Index Scoring in Ulcerative Colitis Using Foundation Models

DGX agent

arXiv:2604.23706v1 Announce Type: new Abstract: Histologic assessment of ulcerative colitis (UC) activity is an important endpoint in clinical trials and routine care, but manual grading with indices

researcharxiv-cs-cv
28 Apr 2026
Applications

When Silence Matters: The Impact of Irrelevant Audio on Text Reasoning in Large Audio-Language Models

DGX agent

arXiv:2510.00626v3 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) unify speech and text processing, but their robustness in noisy real-world settings remains underexplored.

applicationsarxiv-cs-cl
28 Apr 2026
Safety

Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding

DGX agent

arXiv:2512.22519v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have made impressive progress toward general-purpose robotic manipulation by post-training large Vision-L

safetyarxiv-cs-ro
27 Apr 2026
Agents

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale

DGX agent

arXiv:2604.00813v3 Announce Type: replace-cross Abstract: End-to-end autonomous driving has evolved from the conventional paradigm based on sparse perception into vision-language-action (VLA) models,

agentsarxiv-cs-ai
27 Apr 2026
Research

Initial results of the Digital Consciousness Model

DGX agent

arXiv:2601.17060v2 Announce Type: replace-cross Abstract: Artificially intelligent systems have become remarkably sophisticated. They hold conversations, write essays, and seem to understand context i

researcharxiv-cs-ai
27 Apr 2026
Model Releases

LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines

DGX agent

arXiv:2411.08027v3 Announce Type: replace-cross Abstract: Most learning-based approaches to complex physical reasoning sidestep the crucial problem of parameter identification (e.g., mass, friction) t

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Preference Heads in Large Language Models: A Mechanistic Framework for Interpretable Personalization

DGX agent

arXiv:2604.22345v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong implicit personalization ability, yet most existing approaches treat this behavior as a black box, relying o

researcharxiv-cs-cl
27 Apr 2026
Research

Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient Partitioning

DGX agent

arXiv:2604.22072v1 Announce Type: cross Abstract: Federated learning (FL) aggregation on serverless platforms faces a hard scalability ceiling: existing architectures (lambda-FL, LIFL) partition clien

researcharxiv-cs-ai
27 Apr 2026
Agents

Source-Modality Monitoring in Vision-Language Models

DGX agent

arXiv:2604.22038v1 Announce Type: new Abstract: We define and investigate source-modality monitoring -- the ability of multimodal models to track and communicate the input source from which pieces of

agentsarxiv-cs-cl
27 Apr 2026
Model Releases

Sum-of-Checks: Structured Reasoning for Surgical Safety with Large Vision-Language Models

DGX agent

arXiv:2604.22156v1 Announce Type: cross Abstract: Purpose: Accurate assessment of the Critical View of Safety (CVS) during laparoscopic cholecystectomy is essential to prevent bile duct injury, a comp

model-releasesarxiv-cs-cv
27 Apr 2026
Safety

Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models

DGX agent

arXiv:2510.11586v2 Announce Type: replace Abstract: Many in-silico simulations of human survey responses with large language models (LLMs) focus on generating closed-ended survey responses, whereas LL

safetyarxiv-cs-cl
27 Apr 2026
Safety

Transferable Physical-World Adversarial Patches Against Pedestrian Detection Models

DGX agent

arXiv:2604.22552v1 Announce Type: new Abstract: Physical adversarial patch attacks critically threaten pedestrian detection, causing surveillance and autonomous driving systems to miss pedestrians and

safetyarxiv-cs-cv
27 Apr 2026
Research

UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions

DGX agent

arXiv:2604.22209v1 Announce Type: cross Abstract: Generative audio modeling has largely been fragmented into specialized tasks, text-to-speech (TTS), text-to-music (TTM), and text-to-audio (TTA), each

researcharxiv-cs-ai
27 Apr 2026
Research

A Systematic Review and Taxonomy of Reinforcement Learning-Model Predictive Control Integration for Linear Systems

DGX agent

arXiv:2604.21030v1 Announce Type: cross Abstract: The integration of Model Predictive Control (MPC) and Reinforcement Learning (RL) has emerged as a promising paradigm for constrained decision-making

researcharxiv-cs-ai
24 Apr 2026
Safety

Align Generative Artificial Intelligence with Human Preferences: A Novel Large Language Model Fine-Tuning Method for Online Review Management

DGX agent

arXiv:2604.21209v1 Announce Type: new Abstract: Online reviews have played a pivotal role in consumers' decision-making processes. Existing research has highlighted the significant impact of manageria

safetyarxiv-cs-ai
24 Apr 2026
Research

Evaluation of Automatic Speech Recognition Using Generative Large Language Models

DGX agent

arXiv:2604.21928v1 Announce Type: new Abstract: Automatic Speech Recognition (ASR) is traditionally evaluated using Word Error Rate (WER), a metric that is insensitive to meaning. Embedding-based sema

researcharxiv-cs-cl
24 Apr 2026
Research

How Much Is One Recurrence Worth? Iso-Depth Scaling Laws for Looped Language Models

DGX agent

arXiv:2604.21106v1 Announce Type: cross Abstract: We measure how much one extra recurrence is worth to a looped (depth-recurrent) language model, in equivalent unique parameters. From an iso-depth swe

researcharxiv-cs-cl
24 Apr 2026
Research

PercHead: Perceptual Head Model for Single-Image 3D Head Reconstruction & Editing

DGX agent

arXiv:2511.02777v2 Announce Type: replace Abstract: We present PercHead, a model for single-image 3D head reconstruction and disentangled 3D editing - two tasks that are inherently challenging due to

researcharxiv-cs-cv
24 Apr 2026
Safety

Ramen: Robust Test-Time Adaptation of Vision-Language Models with Active Sample Selection

DGX agent

arXiv:2604.21728v1 Announce Type: new Abstract: Pretrained vision-language models such as CLIP exhibit strong zero-shot generalization but remain sensitive to distribution shifts. Test-time adaptation

safetyarxiv-cs-cv
24 Apr 2026
← Previous
1…140141142143144…1030
Next →