AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,811 results
Research

Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow

DGX agent

arXiv:2605.21980v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) represent a significant leap towards empathetic agents, demonstrating remarkable capabilities in emotion understand

researcharxiv-cs-cv
22 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Texts

DGX agent

arXiv:2605.22641v1 Announce Type: new Abstract: Detecting Schwartz values in political text is difficult because implicit cues often depend on surrounding arguments and fine-grained distinctions betwe

researcharxiv-cs-cl
22 May 2026
Research

PolycubeNet: A Dual-latent Diffusion Model for Polycube-Based Hexahedral Mesh Generation

DGX agent

arXiv:2605.20274v1 Announce Type: cross Abstract: Hexahedral meshes are widely used in simulation pipelines, yet automatic generation remains challenging for complex CAD geometries. Polycube-based hex

researcharxiv-cs-ai
22 May 2026
Research

AI-Augmented Surveys: Leveraging Large Language Models and Surveys for Opinion Prediction

DGX agent

arXiv:2305.09620v4 Announce Type: replace Abstract: Nationally representative surveys track public opinion, yet they ask only a limited set of questions each year, limiting its potential to capture hi

researcharxiv-cs-cl
21 May 2026
Applications

Bridging Language Models and Financial Analysis

DGX agent

arXiv:2503.22693v2 Announce Type: replace-cross Abstract: The rapid advancements in Large Language Models (LLMs) have unlocked transformative possibilities in natural language processing, particularly

applicationsarxiv-cs-cl
21 May 2026
Applications

Causal Discovery from Heteroscedastic Stochastic Dynamical Systems under Imperfect Physical Models

DGX agent

arXiv:2602.04907v2 Announce Type: replace Abstract: Causal discovery is a data-driven paradigm for analyzing complex systems, while physics-based models, such as ordinary differential equations (ODEs)

applicationsarxiv-cs-lg
21 May 2026
Research

Conflict-Aware Additive Guidance for Flow Models under Compositional Rewards

DGX agent

arXiv:2605.20758v1 Announce Type: cross Abstract: Inference-time guided sampling steers state-of-the-art diffusion and flow models without fine-tuning by interpreting the generation process as a contr

researcharxiv-cs-cv
21 May 2026
Safety

Correcting Stochastic Update Bias in Preconditioned Language Model Optimizers

DGX agent

arXiv:2605.20756v1 Announce Type: new Abstract: Preconditioned optimizers are central to language model training, but their stochastic update rules are usually treated as direct approximations to popu

safetyarxiv-cs-lg
21 May 2026
Safety

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

DGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

safetyarxiv-cs-cl
21 May 2026
Safety

GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents

DGX agent

arXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple

safetyarxiv-cs-lg
21 May 2026
Research

How Many Human Survey Respondents is a Large Language Model Worth? An Uncertainty Quantification Perspective

DGX agent

arXiv:2502.17773v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate survey responses, but synthetic data can be misaligned with the human populatio

researcharxiv-cs-lg
21 May 2026
Research

iReasoner: Trajectory-Aware Intrinsic Reasoning Supervision for Self-Evolving Large Multimodal Models

DGX agent

arXiv:2601.05877v3 Announce Type: replace Abstract: Recent work shows that large multimodal models (LMMs) can self-improve from unlabeled data via self-play and intrinsic feedback. Yet existing self-e

researcharxiv-cs-cl
21 May 2026
Research

Musical Attention Transformer: Music Generation Using a Music-Specific Attention Model

DGX agent

arXiv:2605.21081v1 Announce Type: cross Abstract: This study aims to enhance the quality of music generation using Transformers by incorporating meta-information. While Transformer-based approaches ar

researcharxiv-cs-lg
21 May 2026
Safety

PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction

DGX agent

arXiv:2605.21414v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation by leveraging large pretrained vision-languag

safetyarxiv-cs-cv
21 May 2026
Research

SDM: A Powerful Tool for Evaluating Model Robustness

DGX agent

arXiv:2605.20308v1 Announce Type: new Abstract: Gradient-based attacks are important methods for evaluating model robustness. However, since the proposal of APGD, it has been difficult for such method

researcharxiv-cs-cv
21 May 2026
Safety

Towards Context-Invariant Safety Alignment for Large Language Models

DGX agent

arXiv:2605.20994v1 Announce Type: new Abstract: Preference-based post-training aligns LLMs with human intent, yet safety behavior often remains brittle. A model may refuse a harmful request in a stand

safetyarxiv-cs-cl
21 May 2026
Local Ai

You Don't Need Attention: Gated Convolutional Modeling for Watch-Based Fall Detection

DGX agent

arXiv:2605.20275v1 Announce Type: new Abstract: Existing deep learning approaches for wearable fall detection systems rely on self-attention mechanisms that impose quadratic computational overhead, di

local-aiarxiv-cs-cv
21 May 2026
Local Ai

3D Modeling and Automated Measurement of Concrete Cracks via Segment Anything Refinement and Visual Inertial LiDAR Fusion

DGX agent

arXiv:2501.09203v2 Announce Type: replace Abstract: Visual-Spatial Systems has become increasingly essential in concrete crack inspection. However, existing methods often lacks adaptability to diverse

local-aiarxiv-cs-cv
20 May 2026
Research

Artificial Phantasia: Emergent Mental Imagery in Large Language Models

DGX agent

arXiv:2509.23108v2 Announce Type: replace Abstract: Can visual imagery be driven solely by language? This idea goes against cognitive science's traditional view that visual mental imagery is only poss

researcharxiv-cs-ai
20 May 2026
Local Ai

BabyMamba-HAR: Lightweight Selective State Space Models for Efficient Human Activity Recognition on Resource Constrained Devices

DGX agent

arXiv:2602.09872v2 Announce Type: replace Abstract: Human activity recognition (HAR) on resource constrained devices requires high accuracy across diverse sensor setups. Selective state space models (

local-aiarxiv-cs-cv
20 May 2026
Safety

Boosting Text-to-Image Diffusion Models via Core Token Attention-Based Seed Selection

DGX agent

arXiv:2605.19532v1 Announce Type: new Abstract: Text-to-image diffusion models can synthesize high-quality images, yet the outcome is notoriously sensitive to the random seed: different initial seeds

safetyarxiv-cs-cv
20 May 2026
Research

Breaking Modality Heterogeneity in Low-Bit Quantization for Large Vision-Language Models

DGX agent

arXiv:2605.19929v1 Announce Type: cross Abstract: Low-bit post-training quantization (PTQ) is a pivotal technique for deploying Vision-Language Models (VLMs) on resource-constrained devices. However,

researcharxiv-cs-ai
20 May 2026
Research

Contextualized Visual Personalization in Vision-Language Models

DGX agent

arXiv:2602.03454v3 Announce Type: replace Abstract: Despite recent progress in vision-language models (VLMs), existing approaches often fail to generate personalized responses based on the user's spec

researcharxiv-cs-cv
20 May 2026
Tutorials

Drifting Objectives for Refining Discrete Diffusion Language Models

DGX agent

arXiv:2605.19470v1 Announce Type: new Abstract: Discrete diffusion language models (DDLMs) generate text by iteratively denoising categorical token sequences, while recent drifting methods for continu

tutorialsarxiv-cs-cl
20 May 2026
Safety

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models

DGX agent

arXiv:2410.15362v2 Announce Type: replace-cross Abstract: Aligned Large Language Models (LLMs) have attracted significant attention for their safety, particularly in the context of jailbreak attacks t

safetyarxiv-cs-ai
20 May 2026
Local Ai

FieldFormer: Locality-Aware Transformers for Spatio-Temporal Modeling on Sparse Sensor Networks

DGX agent

arXiv:2510.03589v2 Announce Type: replace Abstract: Spatio-temporal sensor data in real-world systems is often sparse, noisy, and irregular, making latent field reconstruction fundamentally underconst

local-aiarxiv-cs-lg
20 May 2026
Applications

FlyMirage: A Fully Automated Generation Pipeline for Diverse and Scalable UAV Flight Data via Generative World Model

DGX agent

arXiv:2605.19600v1 Announce Type: new Abstract: In the field of Vision-Language Navigation (VLN), aerial datasets remain limited in their ability to combine scale, diversity, and realism, often relyin

applicationsarxiv-cs-ro
20 May 2026
Research

GenAI-FDIA: Physics-Informed Generative Models for False Data Injection Attacks

DGX agent

arXiv:2605.18873v1 Announce Type: cross Abstract: Training and evaluating false data injection attack (FDIA) detectors for power systems is constrained by data scarcity. Operational grid measurements

researcharxiv-cs-ai
20 May 2026
Safety

Measuring Stereotype and Deviation Biases in Large Language Models

DGX agent

arXiv:2508.06649v3 Announce Type: replace Abstract: Large language models (LLMs) are widely applied across diverse domains, raising concerns about their limitations and potential risks. In this study,

safetyarxiv-cs-cl
20 May 2026
Applications

Position: Graph Condensation Needs a Reset -- Move Beyond Full-dataset Training and Model-Dependence

DGX agent

arXiv:2605.18893v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) are powerful tools for learning from graph-structured data, but their scalability is increasingly strained by the size of r

applicationsarxiv-cs-lg
20 May 2026
Safety

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

DGX agent

arXiv:2605.19940v1 Announce Type: new Abstract: Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cum

safetyarxiv-cs-ai
20 May 2026
Research

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference

DGX agent

arXiv:2605.19218v1 Announce Type: cross Abstract: Vision-Language Models suffer severe KV cache pressure at inference, as a single image often encodes into thousands of tokens. Most existing methods e

researcharxiv-cs-ai
20 May 2026
Tutorials

Towards Fine-Grained Robustness: Attention-Guided Test-Time Prompt Tuning for Vision-Language Models

DGX agent

arXiv:2605.19956v1 Announce Type: new Abstract: Vision-Language Models (VLMs), such as CLIP, have achieved significant zero-shot performance on downstream tasks with various fine-tuning adaptation met

tutorialsarxiv-cs-cv
20 May 2026
Local Ai

Axial-Relation Guided Fusion State Space Model for Optical-Elevation Sensing Image Segmentation

DGX agent

arXiv:2605.16768v1 Announce Type: new Abstract: Semantic segmentation of multi-source remote sensing images is a fundamental task for Earth observation applications. Existing methods often struggle wi

local-aiarxiv-cs-cv
19 May 2026
Local Ai

CanViT: Toward Active-Vision Foundation Models

DGX agent

arXiv:2603.22570v2 Announce Type: replace Abstract: Active computer vision promises efficient, biologically plausible perception through sequential, localized glimpses, but lacks scalable general-purp

local-aiarxiv-cs-cv
19 May 2026
Tutorials

CoLLM-NAS: Collaborative Large Language Models for Efficient Knowledge-Guided Neural Architecture Search

DGX agent

arXiv:2509.26037v2 Announce Type: replace Abstract: The integration of Large Language Models (LLMs) with Neural Architecture Search (NAS) has introduced new possibilities for automating the design of

tutorialsarxiv-cs-ai
19 May 2026
Research

Confidence Geometry Reveals Trace-Level Correctness in Large Language Model Reasoning

DGX agent

arXiv:2605.16824v1 Announce Type: cross Abstract: Large language models (LLMs) generate not only reasoning text, but also token-level confidence trajectories that record how uncertainty evolves during

researcharxiv-cs-cl
19 May 2026
Research

Counterparty Modeling is Not Strategy: The Limits of LLM Negotiators

DGX agent

arXiv:2605.16575v1 Announce Type: new Abstract: Negotiation requires more than inferring what the other side wants: it requires using that information to make advantageous offers and counteroffers ove

researcharxiv-cs-ai
19 May 2026
Research

Does Your Reasoning Model Implicitly Know When to Stop Thinking?

DGX agent

arXiv:2602.08354v5 Announce Type: replace Abstract: Recent advancements in large reasoning models (LRMs) have greatly improved their capabilities on complex reasoning tasks through Long Chains of Thou

researcharxiv-cs-ai
19 May 2026
Agents

DriveMoE: Mixture-of-Experts for Vision-Language-Action Model in End-to-End Autonomous Driving

DGX agent

arXiv:2505.16278v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving (E2E-AD) demands effective processing of multi-view sensory data and robust handling of diverse and complex driv

agentsarxiv-cs-ai
19 May 2026
Research

Embracing Anisotropy: Turning Massive Activations into Interpretable Control Knobs for Large Language Models

DGX agent

arXiv:2603.00029v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit highly anisotropic internal representations, often characterized by massive activations, a phenomenon where a s

researcharxiv-cs-cl
19 May 2026
Research

Finding Sense in Nonsense with Generated Contexts: Perspectives from Humans and Language Models

DGX agent

arXiv:2602.11699v3 Announce Type: replace Abstract: Nonsensical and anomalous sentences have been instrumental in the development of computational models of semantic interpretation. A core challenge i

researcharxiv-cs-cl
19 May 2026
Research

Flow-Direct: Feedback-Efficient and Reusable Guidance for Flow Models via Non-Parametric Guidance Field

DGX agent

arXiv:2605.16348v1 Announce Type: cross Abstract: Training-free guidance enables pre-trained diffusion and flow models to optimize application-specific objectives using feedback from external black-bo

researcharxiv-cs-ai
19 May 2026
Safety

Forgetting is Competition: Rethinking Unlearning as Representation Interference in Diffusion Models

DGX agent

arXiv:2603.00975v2 Announce Type: replace-cross Abstract: Deployed text-to-image diffusion models increasingly require post-hoc concept unlearning for copyright claims, artist opt-outs, safety updates

safetyarxiv-cs-ai
19 May 2026
Safety

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes

DGX agent

arXiv:2605.16303v1 Announce Type: cross Abstract: Large language models (LLM) agents may offer tools to predict human responses to surveys. A common technique for defining these agents uses only demog

safetyarxiv-cs-ai
19 May 2026
Research

Language Acquisition Device in Large Language Models

DGX agent

arXiv:2605.16758v1 Announce Type: new Abstract: Large Language Models (LLMs) remain substantially less data-efficient than humans. Pre-pretraining (PPT) on synthetic languages has been proposed to clo

researcharxiv-cs-cl
19 May 2026
Tutorials

Learning What Evaluators Value: A Reliable Approach to Modeling Evaluator Preferences

DGX agent

arXiv:2605.16615v1 Announce Type: new Abstract: In many applications, human and LLM evaluators use assessments of relevant criteria to create an overall evaluation for an item or individual. For examp

tutorialsarxiv-cs-lg
19 May 2026
Research

Leveraging Graph Structure in Seq2Seq Models for Knowledge Graph Link Prediction

DGX agent

arXiv:2605.18211v1 Announce Type: cross Abstract: We introduce Graph-Augmented Sequence-to-Sequence (GA-S2S), a novel framework that integrates a T5-small encoder-decoder with a Relational Graph Atten

researcharxiv-cs-ai
19 May 2026
← Previous
1…216217218219220…1038
Next →