AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 Jul 2026

No Reliable Evidence of Self-Reported Sentience in Small Large Language Models

Model ReleasesDGX agent

arXiv:2601.15334v2 Announce Type: replace-cross Abstract: Whether language models possess sentience has no empirical answer. But whether they believe themselves to be sentient can, in principle, be te

No Time Like the Present: Agentic Test-Time Training for LLM Agents

SafetyDGX agent

arXiv:2607.03441v1 Announce Type: cross Abstract: LLM agents often degrade over long episodes: as trajectories grow, they revisit explored states, repeat failed actions, and lose strategies that previ

Noisy-Channel Minimum Bayes Risk Decoding

ResearchDGX agent

arXiv:2607.05198v1 Announce Type: cross Abstract: Minimum Bayes Risk (MBR) decoding yields more robust and higher-quality text generation than maximum a posteriori (MAP) decoding by selecting hypothes


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

NormWorlds-CF: Solver-Verified Counterfactual Normative Reasoning with Metamorphic-Relation GRPO

Model ReleasesDGX agent

arXiv:2607.03957v1 Announce Type: cross Abstract: Language models can reach the right normative verdict for the wrong reason. We introduce NormWorlds-CF, a solver-verified environment for counterfactu

Not All Refusals Are Equal: How Safety Alignment Fails Cybersecurity at Scale

Model ReleasesDGX agent

arXiv:2607.02714v1 Announce Type: cross Abstract: There is no doubt that safety alignment is an essential step in LLM training. However, conceptually it does not distinguish between various domains an

Not Every Sync Is Safe: Calibrated DiLoCo Scheduling for Shared AI Infrastructure

Local AiDGX agent

arXiv:2607.02544v1 Announce Type: cross Abstract: DiLoCo-style training reduces communication by letting learner islands train locally before occasional outer synchronization, making it attractive for

NouveauVoice: Generating Novel Pseudo Speakers for Voice Anonymization

ResearchDGX agent

arXiv:2607.03985v1 Announce Type: cross Abstract: Advanced neural technologies in speech synthesis and voice conversion (VC) have introduced severe risks to personal privacy, necessitating robust Spea

NRT-Bench: Benchmarking Multi-Turn Red-Teaming of LLM Operator Agents in Safety-Critical Control Rooms

Model ReleasesDGX agent

arXiv:2606.20408v3 Announce Type: replace-cross Abstract: Large language model (LLM) agents are increasingly proposed as supervisory components for safety-critical systems, yet their robustness under

Obey, Diverge, Collapse: Blind Obedience to Incorrect Instructions Drives Code LLMs to Irrecoverable Code Semantic Collapse

Model ReleasesDGX agent

arXiv:2607.04537v1 Announce Type: cross Abstract: Code language models are now trusted collaborators in production workflows for debugging, refactoring, and iterative repair, and every benchmark that

Object-Centric Environment Modeling for Agentic Tasks

Local AiDGX agent

arXiv:2607.02846v1 Announce Type: new Abstract: Large language model (LLM) agents can improve through accumulated experience, but free-form textual memories become difficult to maintain, validate, and

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

HardwareDGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

OmniFocus: Query-Guided Modality-Balanced Token Compression for Omni-Modal Large Language Models

Model ReleasesDGX agent

arXiv:2607.03050v1 Announce Type: cross Abstract: Omni modal large language models (OmniLLMs) have attracted wide attention for their ability to jointly process audio and video, but they generate larg

OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers

Model ReleasesDGX agent

arXiv:2607.04033v1 Announce Type: cross Abstract: Optimizer selection for large-scale model training has become a system-level design decision constrained jointly by compute, memory, tuning budget, an

OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies

SafetyDGX agent

arXiv:2607.03723v1 Announce Type: cross Abstract: Visual policies learned from human videos, teleoperation, and robot demonstrations offer scalable motion priors, but often fail in contact-rich manipu

On Pairwise Quantile Regression -- Statistical Guarantees and Applications

ResearchDGX agent

arXiv:2607.04431v1 Announce Type: cross Abstract: Quantile regression provides a powerful tool for summarizing the conditional distribution of a real valued random variable (r.v.) of interest Y as a f

On the Ability of Transformers to Verify Plans

TutorialsDGX agent

arXiv:2603.19954v2 Announce Type: replace Abstract: Transformers have shown inconsistent success in AI planning tasks, and theoretical understanding of when generalization should be expected has been

One Framework for All: Cross-Modal Membership Inference for Generative Models

ApplicationsDGX agent

arXiv:2607.04339v1 Announce Type: cross Abstract: Large generative models across text-to-text, text-to-image, and image-to-text modalities have been shown to pose significant privacy risks. One fundam

One Prompt, Many Sounds: Modeling Listener Variability in LLM-Based Equalization

Model ReleasesDGX agent

arXiv:2601.09448v3 Announce Type: replace-cross Abstract: Conventional audio equalization is a static process that requires manual and cumbersome adjustments to adapt to changing listening contexts (e

Online Linear Programming for Multi-Objective Routing in LLM Serving

SafetyDGX agent

arXiv:2607.03948v1 Announce Type: new Abstract: We study the online routing problem in large language model serving, where requests arrive sequentially and must be dispatched to parallel decode worker

Open Problems in AI Incident Governance

SafetyDGX agent

arXiv:2607.05163v1 Announce Type: cross Abstract: AI systems may produce failures after deployment that pre-deployment safety assessments do not anticipate. Managing these failures requires what we re

OpenGlass: A Sensing-Computing Split Architecture for Local MLLM-Driven Real-Time Visual Assistance

Local AiDGX agent

arXiv:2607.03213v1 Announce Type: cross Abstract: We present OpenGlass, an open-source, privacy-oriented, local-first system for low-latency multimodal visual assistance, with a primary focus on blind

OpenTinker: Separating Concerns in Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2601.07376v2 Announce Type: replace Abstract: We introduce extsc{OpenTinker}, an open infrastructure for training large language model (LLM) agents with many LoRA-backed policies over shared exe

Operator-on-F complements value-equivalence: a planning-time diagnostic for latent world models

ResearchDGX agent

arXiv:2607.04464v1 Announce Type: cross Abstract: World-model evaluation for model-based reinforcement learning typically asks whether the learned model predicts reward and value well, which can leave

OptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative Refinement

AgentsDGX agent

arXiv:2607.05346v1 Announce Type: new Abstract: We propose OptiAgent, a multi-agent framework that, given a natural language description of an Operations Research problem, is able to output a solver-r

Optimal-Agent-Selection: State-Aware Routing Framework for Efficient Multi-Agent Collaboration

AgentsDGX agent

arXiv:2511.02200v2 Announce Type: replace Abstract: The emergence of multi-agent systems powered by large language models (LLMs) has unlocked new frontiers in complex task-solving, enabling diverse ag

Optimizing ML Workload Partitioning between CPUs and CIM Accelerators for Heterogeneous Computing

ResearchDGX agent

arXiv:2607.05240v1 Announce Type: cross Abstract: Computing-in-Memory (CIM) accelerators execute Matrix-Vector Multiplications (MVMs) in memory, making them a compelling solution for Machine Learning

Order-based Causal Discovery for Multistage Processes

ResearchDGX agent

arXiv:2607.03971v1 Announce Type: cross Abstract: Causality has become an increasingly important tool for gaining a deeper understanding of complex systems. Among various causal analysis methods, caus

Organizational Memory for Agentic Business Process Execution

AgentsDGX agent

arXiv:2607.03228v1 Announce Type: new Abstract: LLM-based agents offer new opportunities for automating business process execution beyond the limits of rule-based systems. However, general-purpose LLM

OrthoReg: Orthogonal Regularization for Hybrid Symbolic-Neural Dynamical Systems

Model ReleasesDGX agent

arXiv:2606.19145v2 Announce Type: replace-cross Abstract: Dynamical systems are fundamental to modeling the natural world, yet modeling them involves a persistent trade-off: manually prescribed mechan

OSF: On Pre-training and Scaling of Sleep Foundation Models

Model ReleasesDGX agent

arXiv:2603.00190v2 Announce Type: replace-cross Abstract: Polysomnography (PSG) provides the gold standard for sleep assessment but suffers from substantial heterogeneity across recording devices and

Out-of-Distribution Generalization of Risk Aversion in Language Models

Model ReleasesDGX agent

arXiv:2607.02755v1 Announce Type: cross Abstract: Training AIs to be risk-averse in resources could offer a failsafe in the event that AIs turn out misaligned. Misaligned but risk-averse AIs would ten

Oyster-II: Reinforcement Learning for Constructive Safety Alignment in Large Language Models

SafetyDGX agent

arXiv:2607.02914v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated remarkable capabilities across diverse applications, yet ensuring their simultaneous safety, helpfulness,

Panorama: Fast-Track Nearest Neighbors

ResearchDGX agent

arXiv:2510.00566v4 Announce Type: replace-cross Abstract: Approximate Nearest-Neighbor Search (ANNS) pipelines for high-dimensional neural embeddings spend the bulk of their query time in candidate ve

Parallelized Autoregressive Decoding for Omni-Modal Dense Video Captioning

ResearchDGX agent

arXiv:2607.02963v1 Announce Type: cross Abstract: Dense video captioning aims to generate temporally grounded descriptions of video events, benefiting both event-level video understanding and generati

Parameter Efficient Multimodal Instruction Tuning for Romanian Vision Language Models

Model ReleasesDGX agent

arXiv:2512.14926v2 Announce Type: replace-cross Abstract: Focusing on low-resource languages is an essential step toward democratizing generative AI. In this work, we contribute to reducing the multim

Parametric Memory Decoding for Zero-Shot Routing in LoRA-Based External Parametric Memory

TutorialsDGX agent

arXiv:2607.04118v1 Announce Type: cross Abstract: With the rise of parametric memory, LoRA-based External Parametric Memory (EPM) has emerged as a modular solution, but existing routing methods often

Parity-Aware Byte-Pair Encoding: Improving Cross-lingual Fairness in Tokenization

SafetyDGX agent

arXiv:2508.04796v3 Announce Type: replace-cross Abstract: Tokenization is the first -- and often least scrutinized -- step of most NLP pipelines. Standard algorithms for learning tokenizers rely on fr

PDEFlow: Autonomous Agentic PDE Pipelines for Neural Operator Learning and Solver-Free Inference

Model ReleasesDGX agent

arXiv:2607.05134v1 Announce Type: cross Abstract: We present PDEFlow, an autonomous agentic framework that turns user-level ODE and PDE descriptions into solver-backed neural-operator pipelines. The w

PDFBench: A Benchmark for De novo Protein Design from Function

Model ReleasesDGX agent

arXiv:2505.20346v3 Announce Type: replace-cross Abstract: Function-guided protein design is a crucial task with significant applications in drug discovery and enzyme engineering. However, the field la

PedestrianDiffusion: Multimodal Generative Denoising and Dense State Estimation for Inertial Navigation

ApplicationsDGX agent

arXiv:2607.03349v1 Announce Type: cross Abstract: The accuracy of consumer-grade inertial navigation is bottlenecked by the stochastic noise of Micro-Electro-Mechanical Systems (MEMS). Traditional det

PEEK: Predictive Queue-Informed KV Cache Management for LLM Serving

HardwareDGX agent

arXiv:2607.02525v1 Announce Type: cross Abstract: We present PEEK, a lightweight scheduling and eviction framework for both online (streaming) and offline (batch) LLM serving; this paper focuses on th

Personalized Causal Recourse: A Human-In-The-Loop Approach

ResearchDGX agent

arXiv:2607.03425v1 Announce Type: new Abstract: Algorithmic recourse addresses the challenge of providing tailored recommendations to users affected by unfavorable machine learning decisions, in poten

pFedNavi: Structure-Aware Personalized Federated Vision-Language Navigation for Embodied AI

Model ReleasesDGX agent

arXiv:2602.14401v2 Announce Type: replace-cross Abstract: Vision-Language Navigation VLN requires large-scale trajectory instruction data from private indoor environments, raising significant privacy

Phase-Preserving Trimodal Transformer for Tropical Forest Biomass Estimation Using Optical and PolInSAR Data

Local AiDGX agent

arXiv:2607.03663v1 Announce Type: cross Abstract: The accurate estimation of Above-Ground Biomass (AGB) in mature tropical forests remains a critical challenge in remote sensing, primarily due to the

Piercing Gilbreath's Conjecture: From Deep Number Theory Insights to Fintech and Cybersecurity

ResearchDGX agent

arXiv:2607.04166v1 Announce Type: cross Abstract: I propose a new methodology to attack the fascinating Gilbreath's conjecture about prime numbers, first posted in 1878 and unsolved to this day. The p

PixCon: Clean-Positive Contrastive Learning for Foundation-Model Semi-Supervised Segmentation

ResearchDGX agent

arXiv:2607.03068v1 Announce Type: cross Abstract: Semi-supervised semantic segmentation (SSSS) has long turned on one question, which pseudo-labels to trust, and answered it with ever more careful con

PLACEMEM: Toward a Compute-Aware Memory Plane for Lifelong Agents

Model ReleasesDGX agent

arXiv:2607.04089v1 Announce Type: new Abstract: Lifelong agents need more than larger context windows and better retrieval. They need memories that can persist, evolve, and be corrected without forcin

PLGSA-Transformer: Periocular Landmark-Guided Attention with Occlusion-Adaptive Cosine Thresholding for Cross-Modal Masked and Unmasked Face Recognition

ApplicationsDGX agent

arXiv:2607.03581v1 Announce Type: cross Abstract: The widespread adoption of facial masks, accelerated by COVID-19 and mandated in security-sensitive settings, has exposed limitations of conventional

Polarity Detection of Sustainable Development Goals in News Text

Model ReleasesDGX agent

arXiv:2509.19833v4 Announce Type: replace-cross Abstract: The United Nations' Sustainable Development Goals (SDGs) provide a globally recognised framework for addressing major societal, environmental,

Policy Improvement with Style-Specific Demonstrations

SafetyDGX agent

arXiv:2506.16995v4 Announce Type: replace Abstract: Proficient game agents with diverse play styles enrich the gaming experience and enhance the replay value of games. However, recent advancements in

Pooling-Based Context Modeling for Convolution-Free Deep Image Prior

Model ReleasesDGX agent

arXiv:2607.02952v1 Announce Type: cross Abstract: Convolutional Neural Networks (CNNs) achieve strong denoising performance by exploiting spatial context from neighboring pixels. Deep Image Prior (DIP

Position: Use Sparse Autoencoders to Discover Unknowns

SafetyDGX agent

arXiv:2506.23845v2 Announce Type: replace-cross Abstract: While sparse autoencoders (SAEs) have generated significant excitement, a series of negative results have added to skepticism about their usef

Post-Generation Curation of Synthetic Images via Homogeneous-Heterogeneous Splitting

SafetyDGX agent

arXiv:2607.02637v1 Announce Type: cross Abstract: Recent generative models can produce high-quality synthetic images, offering scalable training training data for data-hungry models. Existing approach

PosterHarness: Turning Scientific Poster Generation into an Auditable Instruction-Following Benchmark

Model ReleasesDGX agent

arXiv:2607.03006v1 Announce Type: cross Abstract: Text-rich image models can now design poster-scale layouts, but we lack ways to measure whether they honor scientific communication contracts: legible

PotatoGANs: Utilizing Generative Adversarial Networks, Instance Segmentation, and Explainable AI for Enhanced Potato Disease Identification and Classification

ResearchDGX agent

arXiv:2405.07332v2 Announce Type: cross Abstract: Numerous applications have resulted from the automation of agricultural disease segmentation using deep learning techniques. However, when applied to

PPE-Bench: A Benchmark for Evaluating MLLM Unlearning under Private-Public Entanglement

Model ReleasesDGX agent

arXiv:2607.02897v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong capabilities, but they may memorize private information from web data, raising privacy conc

Predicting Biased Human Decision-Making with Large Language Models in Conversational Settings

Model ReleasesDGX agent

arXiv:2601.11049v2 Announce Type: replace-cross Abstract: We examine whether large language models (LLMs) can predict biased decision-making in conversational settings, and whether their predictions c

Predicting Drafted Deck Strength for 'Magic: the Gathering'

Model ReleasesDGX agent

arXiv:2607.04782v1 Announce Type: cross Abstract: Many real-world games do not admit a fixed, compact rule set: instead, their dynamics are defined by interactions among a large and often evolving col

Predicting Therapeutic Outcome via Aligning Patient-Specific Knowledge Graph and Gene-Level Perturbation Representations

ResearchDGX agent

arXiv:2607.04557v1 Announce Type: cross Abstract: Accurate prediction of patient-specific therapeutic response from pre-treatment transcriptomes is hindered by the scarcity of matched clinical respons

Pretraining Curricula Enable Selective Fine-tuning

SafetyDGX agent

arXiv:2607.04846v1 Announce Type: cross Abstract: Transformers follow implicit curricula whereby some tasks are learned before others. However, how explicit pretraining curricula influence learning, g

← Previous
1…8990919293…358
Next →