AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
29 Jul 2026

Matryoshka Agent: Unfolding Sub-Agents for Long-Horizon Machine Learning Engineering

AgentsDGX agent

arXiv:2607.25090v1 Announce Type: new Abstract: Machine learning engineering (MLE) tasks require long-horizon decision making over iterative solution debugging and refinement, under expensive and feed

MDTransformer: A Hardware-Software Co-Design of Mode-Division Photonic Transformer Accelerator with Inverse-Designed Coherent Crossbar

HardwareDGX agent

arXiv:2607.26016v1 Announce Type: cross Abstract: Recently, photonic transformer accelerators (PTAs) have successfully achieved significant speedup and energy efficiency improvements over electronic a

Measuring and Improving Behavioral Consistency in Large Language Models through Fact-Heuristic-Emotion State Enforcement

Research

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.24765v1 Announce Type: cross Abstract: Large language models (LLMs) can give different answers to the same decision problem across runs, and reverse a decision when their own prior answer r

Measuring the State of Open Science in Transportation Using Large Language Models

ResearchDGX agent

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their p

Mechanisms of Width Scaling in Normalized Residual Networks: The Effective Alignment Dimension

Model ReleasesDGX agent

arXiv:2607.24887v1 Announce Type: cross Abstract: Existing theories of neural-network width characterize asymptotic limits, but provide limited guidance on whether an expansion direction identified fr

MedJudgeRAG: Option-Wise Evidence Judgment with Dynamic Knowledge Graphs for Medical MCQA

Model ReleasesDGX agent

arXiv:2607.24838v1 Announce Type: cross Abstract: In medical multiple-choice question answering (MCQA), Retrieval-Augmented Generation (RAG) can supplement the domain knowledge of language models (LMs

MemLens: A Value-Aware Memory Management System with Interactive Analytics for LLM-based Agents

ResearchDGX agent

arXiv:2607.25992v1 Announce Type: cross Abstract: Recently, memory management has become a key infrastructure for LLM-based agents, as it directly affects long-horizon reasoning, personalized response

Messier: A High-Resolution Corpus for Cross-Benchmark Agent Evaluation

Model ReleasesDGX agent

arXiv:2607.25891v1 Announce Type: new Abstract: Evaluating AI agents in interactive environments is hindered by fragmented tasks, scaffolds, verifiers, and scoring rules. Existing efforts focus on nar

Minimizing Targeted Activations: Input-Only Suppression of Evaluation-Awareness Latents in Large Language Models

Model ReleasesDGX agent

arXiv:2607.25907v1 Announce Type: cross Abstract: Activation steering controls model behavior by editing internal activations at inference time. We study its input-side dual: optimizing a fluent promp

MODUS: Decoder-Only Any-to-Any Modeling of Diverse Modalities

ResearchDGX agent

arXiv:2607.25948v1 Announce Type: cross Abstract: Any-to-any models predict any modality from any combination of others within a single network, a formulation used in multimodal vision and vision-lang

Multi-Scale Structural Features for Continual, Comprehensible Visual Recognition in a Developmental Learning Framework

Model ReleasesDGX agent

arXiv:2607.25531v1 Announce Type: cross Abstract: Contemporary machine learning struggles to learn continually, reuse prior knowledge, and expose a comprehensible internal structure. A recently propos

Multi-Sensor Alignment for Weather Simulations

SafetyDGX agent

arXiv:2607.25612v1 Announce Type: new Abstract: Perception tasks for autonomous vehicles need to work satisfactorily in adverse weather conditions. Due to lack of real-world weather datasets, weather

Multiclass Classification without Labels via Posterior Simplex Geometry

ResearchDGX agent

arXiv:2607.24943v1 Announce Type: cross Abstract: In many classification problems, reliable instance-level labels are unavailable. However, it is often possible to construct weakly enriched unlabeled

Multimodal Hybrid Retrieval-Augmented Generation for Scientific Document Understanding using Open-Source SLMs

Model ReleasesDGX agent

arXiv:2607.24799v1 Announce Type: cross Abstract: Large Language Models tend to hallucinate when answering domain-specific ques tions from scientific documents without prior fine-tuning. Currently, me

MusiChat: Vibe Composing for Music Creation

ResearchDGX agent

arXiv:2607.24873v1 Announce Type: new Abstract: Recent advances in AI music generation have enabled users to create complete musical pieces from natural-language prompts. However, most existing system

MyMentorLLM: A psychotherapy GenAI environment with multimodal voice/text patients, trainees and experts for deliberate practice

ResearchDGX agent

arXiv:2607.25667v1 Announce Type: cross Abstract: Psychotherapists need repeated training and supervision by experts; however, scalability is problematic. Here we present MyMentorLLM, a multimodal voi

Neural Network Learning of One-Bit Protocols for Qubit Measurement Simulation

ResearchDGX agent

arXiv:2607.23645v1 Announce Type: cross Abstract: Communication complexity provides a natural framework for quantifying the classical resources required to reproduce quantum statistics. In the qubit p

Nudging Sustainable Choices through LLM-Generated Recommendation Explanations

ResearchDGX agent

arXiv:2607.25726v1 Announce Type: new Abstract: Recommender systems mediate everyday consumption, offering a promising channel for encouraging sustainable choices. Prior research shows that explanatio

Observing sycophantic AI validate others reduces its appeal but not its persuasiveness

ResearchDGX agent

arXiv:2607.25166v1 Announce Type: new Abstract: AI chatbots can be ``sycophantic,'' or overly agreeable and flattering toward users. Sycophantic AI has been shown to entrench attitudes, yet users freq

ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning

SafetyDGX agent

arXiv:2607.25369v1 Announce Type: new Abstract: Agentic systems have rapidly advanced in their ability to interact with real-world environments, leverage external tools, and provide services for users

OmniDelta: Skill-Driven Budget Allocation for Token Compression in OmniLLMs

Local AiDGX agent

arXiv:2607.25669v1 Announce Type: new Abstract: Emerging Omni-modal Large Language Models (OmniLLMs) enable unified understanding of text, audio, and video, but their long audio-video token sequences

OmniPhys: Knowledge-Graph-Driven Benchmarking and Collective Optimization for Physical Commonsense in Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2607.25641v1 Announce Type: cross Abstract: While text-to-image models exhibit remarkable visual fidelity, they frequently violate fundamental physical commonsense. Existing benchmarks often rel

OmniQEC: discovering practical quantum error-correcting codes by an AI scientist

ResearchDGX agent

arXiv:2607.25865v1 Announce Type: cross Abstract: Quantum error correction (QEC) is indispensable for scalable fault-tolerant quantum computing. However, discovering QEC codes that remain effective is

On the Design and Evaluation of Human-centered Explainable AI Systems: A Systematic Review and Taxonomy

TutorialsDGX agent

arXiv:2510.12201v2 Announce Type: replace Abstract: As AI becomes more common in everyday living, there is an increasing demand for intelligent systems that are both performant and understandable. Exp

On the Use of LLMs for Specialised Terminology: A Good Alternative to Corpora?

Model ReleasesDGX agent

arXiv:2607.24784v1 Announce Type: new Abstract: Specialised translation relies on the use of documentary and terminological resources, including corpora. These resources are particularly useful for te

OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image Analysis

SafetyDGX agent

arXiv:2607.25108v1 Announce Type: cross Abstract: Biomedical image analysis spans diverse modalities and tasks, yet real-world deployment is hindered by severe distribution shifts across scanners, pro

OrchBench: Evaluating Multi-Agent Orchestration Plans in Isolation via Deterministic Simulation

Model ReleasesDGX agent

arXiv:2607.25656v1 Announce Type: new Abstract: Complex tasks often decompose into parallelizable yet interdependent subtasks, making orchestration critical to the performance of multi-agent systems (

OrganLens: Organ-Specific Representation Learning for CT Foundation Models

ResearchDGX agent

arXiv:2607.25164v1 Announce Type: cross Abstract: A CT examination captures multiple organs, but many biomedical questions concern abnormalities, prognosis, or longitudinal change in a specific organ.

Pass the Baton: Trajectory-Relayed On-Policy Distillation

Model ReleasesDGX agent

arXiv:2607.26057v1 Announce Type: cross Abstract: On-policy distillation (OPD) grounds token-level supervision in the student's own trajectory, yet suffers from prefix failure: once the student commit

PATHFinder Agent for Tailored Prenatal Care

Model ReleasesDGX agent

arXiv:2607.24768v1 Announce Type: new Abstract: Prenatal care is an important preventive service designed to improve outcomes for pregnant individuals. The American College of Obstetricians and Gyneco

PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents

Model ReleasesDGX agent

arXiv:2607.25485v1 Announce Type: new Abstract: Health AI is evolving from answering questions to agentic systems that converse with patients, reason about health records, and act on their behalf. Pri

Patterns of Learner-AI Interaction and Academic Performance in an Object-Oriented Programming Course

TutorialsDGX agent

arXiv:2607.24755v1 Announce Type: cross Abstract: This full research paper examines how different forms of learner-AI interaction relate to learning outcomes in object-oriented programming (OOP) cours

Penelope: Localized Latent Recurrence for Efficient Structured Reasoning

Model ReleasesDGX agent

arXiv:2607.25915v1 Announce Type: new Abstract: Complex structured reasoning tasks often require additional computation, yet current language models obtain it mainly by increasing parameter scale or b

Personalization, Personas, and Forecasting in Value Alignment

Model ReleasesDGX agent

arXiv:2607.24782v1 Announce Type: new Abstract: LLM behavior may be conditioned by human identity in several ways: they may be asked to adapt to users, role-play populations, or forecast how people wo

Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical-Flow Supervision

Model ReleasesDGX agent

arXiv:2607.25321v1 Announce Type: new Abstract: Video diffusion models generate visually compelling content but routinely violate elementary physics when the subject involves fluids: liquid columns br

Physics-Informed Broad Learning System: An Efficient Backpropagation-Free Framework for Solving Partial Differential Equations

ResearchDGX agent

arXiv:2607.25608v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs) by embedding governing

Physics-Informed Neural Operator for Warm-Starting Background-Decomposed and Preconditioned PSFD: Enabling Scalable 3-D EUV Mask Simulation

ResearchDGX agent

arXiv:2607.25330v1 Announce Type: cross Abstract: We present a physics-informed neural operator (PINO) trained with pseudo-spectral frequency-domain (PSFD) equations for electromagnetic (EM) scatterin

Pictura: Perspective-View Self-Play at Scale for Driving

SafetyDGX agent

arXiv:2607.26005v1 Announce Type: cross Abstract: Self-play in simulation produces robust driving policies at scale. Demonstrations of such behavior have been made using privileged vectorized observat

pimathbf{R}^2: Reactive Real-time Flow Policies

SafetyDGX agent

arXiv:2607.26055v1 Announce Type: cross Abstract: Generalist manipulation policies increasingly take the form of action-chunking flow policies built on large pretrained backbones. Such chunks run open

PLATO: Pointer Learner for Agent and Task Openness

SafetyDGX agent

arXiv:2607.25082v1 Announce Type: new Abstract: Open agent systems (OASYS) are increasingly prevalent in real-world domains where the sets of agents and tasks change unpredictably over time. Such open

PreDiff-LM: Pretrained Discrete Masked Diffusion Language Modeling with Hybrid Attention

ResearchDGX agent

arXiv:2607.25157v1 Announce Type: new Abstract: Discrete masked diffusion language models support bidirectional generation and infilling, but adapting pretrained autoregressive (AR) transformers requi

Preliminary Guidelines for Using and Evaluating GenAI Tools to Support Systematic Literature Reviews

TutorialsDGX agent

arXiv:2607.24991v1 Announce Type: cross Abstract: Context: Generative AI (GenAI) and Large Language Models (LLMs) are increasingly used for academic tasks in software engineering and beyond, including

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop

Local AiDGX agent

arXiv:2607.24770v1 Announce Type: new Abstract: Procedural tasks such as furniture assembly and home repair impose substantial cognitive demands because users must interpret instructions, track task p

Psychological Influences of Conversational AI: Research and Design Directions for Reducing Harm and Promoting Well-Being

ResearchDGX agent

arXiv:2607.25057v1 Announce Type: new Abstract: As conversational AI systems become increasingly integrated into daily life, their potential effects on user well-being require ongoing attention. While

Quotient Dynamics, Effective Curvature, and Implicit Bias in Positive Quadratic Networks

SafetyDGX agent

arXiv:2607.25624v1 Announce Type: new Abstract: Positive quadratic networks admit the low-rank representation f_U(x)=x^top UU^top x, where Uinmathbb{R}^{dtimes r} is identifiable only up to right orth

Rashomon Alignment

SafetyDGX agent

arXiv:2607.25680v1 Announce Type: cross Abstract: We propose Rashomon Alignment (RA), a new measure to assess functional similarity between two models. Existing functional similarity measures are dist

Raven: High-Recall Sequence Modeling with Sparse Memory Routing

ResearchDGX agent

arXiv:2607.25357v1 Announce Type: cross Abstract: Long-context recall in linear-time sequence models highlights a tradeoff in how they write to memory. State-based linear models, such as state-space m

Reading Without a Reader: Large Language Models Collapse Reading and Writing into a Single Entangled Code

ApplicationsDGX agent

arXiv:2607.24797v1 Announce Type: cross Abstract: In the literate human brain, reading and writing are two doubly-dissociable systems: a ventral decoding route (impaired in pure alexia) and a fronto-p

Real-Time Driver Safety Scoring Through Inverse Crash Probability Modeling

SafetyDGX agent

arXiv:2603.14841v3 Announce Type: replace-cross Abstract: Road crashes remain a leading cause of preventable fatalities. Existing prediction models predominantly produce binary outcomes, which offer l

Real-time Spatial Retrieval Augmented Generation for Urban Environments

ResearchDGX agent

arXiv:2505.02271v2 Announce Type: replace Abstract: The proliferation of Generative Artificial Ingelligence (AI), especially Large Language Models, presents transformative opportunities for urban appl

Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding

SafetyDGX agent

arXiv:2607.24794v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate superior generalization in fundamental video tasks, restricted context windows limit their lo

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension

Model ReleasesDGX agent

arXiv:2512.06276v3 Announce Type: replace-cross Abstract: Referring Expression Comprehension (REC) is a vision-language task that localizes a specific image region based on a textual description. Exis

Reinforcement Learning for Code Optimization

Model ReleasesDGX agent

arXiv:2607.25970v1 Announce Type: cross Abstract: RL for code correctness is now established: have the model generate a program, run it against hidden test cases, and reward solutions that pass. Exten

ReLATE: Reliability-Guided Evidence Fusion for Robust UAV--Satellite cross-view Geo-Localization

Model ReleasesDGX agent

arXiv:2607.25524v1 Announce Type: cross Abstract: Unmanned aerial vehicle (UAV)-satellite cross-view geo-localization matches UAV images against satellite imagery and has achieved impressive accuracy

REPREC: Representation Driven Parameter-Efficient Recommendation System

Model ReleasesDGX agent

arXiv:2607.24845v1 Announce Type: cross Abstract: Large language models (LLMs) have been applied to sequential recommendation by formulating it as a natural language task. Previous work has improved p

Rethinking Likelihood distributions: Student's t Likelihood Boosts Bayesian Neural Network Performance

ApplicationsDGX agent

arXiv:2607.25376v1 Announce Type: cross Abstract: In Bayesian neural networks (BNNs), variational inference is a widely adopted framework for modeling uncertainty in a distributional way, with the evi

Retraction-Free Optimization over the Stiefel Manifold for the LoRA Fine-Tuning

Model ReleasesDGX agent

arXiv:2607.25299v1 Announce Type: cross Abstract: Optimization over the Stiefel manifold plays a significant role in various machine learning tasks. Existing methods either use the retraction operator

Retrieval-Augmented Generation in LLMs for Mental Health: Quantifying the Incremental Contribution of Retrieval Within a Layered Safety Architecture

SafetyDGX agent

arXiv:2607.24817v1 Announce Type: cross Abstract: Digital mental health interventions (DMHIs) offer scalable support, but ensuring they accurately detect users' intent during volatile situations can b

RIDGE: An Autonomous Framework for Validation and Method Discovery in LLM-Generated Option Pricing

Model ReleasesDGX agent

arXiv:2607.25199v1 Announce Type: cross Abstract: Automated code generation is becoming an important tool in quantitative finance, where large language models can generate option pricing implementatio

Right-sizing Recommendations (RSR): Cloud Workload Conformal Prediction for Virtual Machines in Data Center Operations

ResearchDGX agent

arXiv:2607.24773v1 Announce Type: new Abstract: Managing cloud infrastructure efficiently, especially in environments of large cloud providers or hyperscalers, requires optimizing the use of physical

← Previous
1…4344454647…354
Next →