AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
6 Aug 2026

Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Mediated Reasoning

AgentsDGX agent

arXiv:2608.04457v1 Announce Type: cross Abstract: As 'AI Scientists' emerge to drive research via the Model Context Protocol (MCP), systems relying on ephemeral scripts will fail. The sheer scale of s

Emergence of Hierarchical Emotion Organization in Large Language Models

ResearchDGX agent

arXiv:2507.10599v3 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly power conversational agents, understanding how they model users' emotional states is critical for

EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and Progressive Alignment

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.04472v1 Announce Type: cross Abstract: The development of foundation models (FMs) is crucial for advancing endoscopic image analysis. However, existing endoscopy FMs mainly rely on self-sup

Equitable System-Prompt Selection via Constrained Mixed-Strategy GroupDRO

ApplicationsDGX agent

arXiv:2608.04339v1 Announce Type: cross Abstract: Large language models are increasingly used for information seeking, yet semantically equivalent questions phrased in different ways can receive answe

EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive Decision Tasks

Model ReleasesDGX agent

arXiv:2608.04549v1 Announce Type: cross Abstract: Frontier LLMs are increasingly put to use on open-ended complex questions, different in nature from the ones they are typically evaluated on. We dedic

EviGraph: Evidence-Guided Autonomous Research Agents

AgentsDGX agent

arXiv:2608.04738v1 Announce Type: new Abstract: Autonomous research agents can generate hypotheses, execute experiments, and draft manuscripts, yet their outputs often contain unsupported claims and i

ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Generation

Model ReleasesDGX agent

arXiv:2608.04439v1 Announce Type: cross Abstract: Large language models (LLMs) have made notable progress in code generation, but they still struggle on challenging tasks that require sophisticated al

Explicit Language Memory for Long-Horizon Planning in Vision-Language-Action Models

ResearchDGX agent

arXiv:2608.04765v1 Announce Type: cross Abstract: Vision-language-action (VLA) models provide a unified paradigm for connecting visual perception, language understanding, and robotic control. However,

FBID: Adaptive Personalized Federated Learning for Robust Out-of-Distribution Attack Detection in IoT Networks

ResearchDGX agent

arXiv:2608.04073v1 Announce Type: cross Abstract: Personalized Federated Learning (PFL) has emerged as a promising solution for intrusion detection in heterogeneous IoT environments, as it can improve

Feedback Loops and Code Perturbations in LLM-based Software Engineering: A Case Study on a C-to-Rust Translation System

SafetyDGX agent

arXiv:2512.02567v2 Announce Type: replace-cross Abstract: The advent of strong generative AI has a considerable impact on various software engineering tasks such as code repair, test generation, or la

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning

SafetyDGX agent

arXiv:2608.04771v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) excel on complex tasks through long chain-of-thought (CoT) reasoning, but their lengthy intermediate steps cause severe ov

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

Model ReleasesDGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

Model ReleasesDGX agent

arXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompt

FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation

Model ReleasesDGX agent

arXiv:2608.04374v1 Announce Type: cross Abstract: Large language models can produce fluent financial analysis, but fluency alone does not establish whether a report is suitable for institutional deliv

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

Model ReleasesDGX agent

arXiv:2511.07322v3 Announce Type: replace-cross Abstract: While LLMs have shown great success in financial tasks like stock prediction and question answering, their application in fully automating Equ

Formal Analysis and Supply Chain Security for Agentic AI Skills

Model ReleasesDGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

From Score Matrices to Football-Aware Match-State Simulation: An Auditable LLM Harness for Exact-Score Reranking

Model ReleasesDGX agent

arXiv:2608.05030v1 Announce Type: new Abstract: Football score forecasting combines a strong statistical core with a difficult contextual edge. Dynamic Poisson-family models estimate team strength, ex

FUSEP: A Multi-Center Benchmark for Diverse Tasks in Early Pregnancy Fetal Ultrasound Screening

Model ReleasesDGX agent

arXiv:2608.04766v1 Announce Type: cross Abstract: A large number of infants with congenital anomalies are born each year globally, especially in areas with underdeveloped medical resources. Currently,

Generative Optimization for Incentivized Advertising with Global Level Constraints

SafetyDGX agent

arXiv:2608.04421v1 Announce Type: cross Abstract: Incentivized advertising allocates monetary or virtual rewards to drive user engagement, where a key challenge is optimizing continuous incentive magn

GeoReward: Mitigating Contextual Variable Overestimation in Vision-Language Models for Cross-Market Preference Prediction

SafetyDGX agent

arXiv:2608.04504v1 Announce Type: cross Abstract: Vision-language models excel in many multimodal tasks but remain prone to a subtle yet impactful failure mode: they tend to overestimate dominant visu

Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming

SafetyDGX agent

arXiv:2608.04018v1 Announce Type: cross Abstract: AI agents are increasingly embedded in organizational workflows, where they interact with external information sources and invoke digital tools to per

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning

Model ReleasesDGX agent

arXiv:2608.05045v1 Announce Type: cross Abstract: Released aligned large language models remain vulnerable to malicious downstream finetuning. Existing defenses are largely designed for the fine-tunin

GRALS: GCN-Guided Redundancy-Aware Local Search for Minimum Vertex Cover

Model ReleasesDGX agent

arXiv:2503.06396v2 Announce Type: replace Abstract: The minimum vertex cover (MVC) problem seeks to identify the smallest set of vertices that cover all edges in an undirected graph. As a fundamental

GUARD: Grounding Uncertainty and Ablation-Based Risk Detection for Diffusion-Based VLAs

Model ReleasesDGX agent

arXiv:2608.04510v1 Announce Type: cross Abstract: Diffusion-based vision-language-action (VLA) policies can generate plausible actions even when their predictions are weakly grounded in the visual and

Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent

SafetyDGX agent

arXiv:2608.04772v1 Announce Type: cross Abstract: Scaling supervision for multi-turn medical agents is difficult because expert dialogue annotation is costly and clinical conversations are privacy-res

Hallucinations on the Board: Tool-Augmented Evaluation of LLM Chess Commentary

Model ReleasesDGX agent

arXiv:2608.04240v1 Announce Type: cross Abstract: Superhuman game engines in domains like chess have made expert-level evaluations easily accessible, yet they communicate what is true without the natu

HALT: Verification-Aware Stopping for Retrieval-Augmented Search Agents

SafetyDGX agent

arXiv:2608.02009v2 Announce Type: replace Abstract: Retrieval-augmented search agents answer multi-hop questions by repeatedly issuing search queries and accumulating evidence. This creates a stopping

Hardware Design and Security in the Era of Chiplets and LLMs

ApplicationsDGX agent

arXiv:2608.05063v1 Announce Type: cross Abstract: The semiconductor industry is undergoing a dual revolution: the shift toward heterogeneous 2.5D chiplet systems and the integration of Large Language

Hierarchical Graph Memory for LLM Agents with Path-level Localization and Rewrite

ResearchDGX agent

arXiv:2608.05095v1 Announce Type: new Abstract: Agents for long term reasoning require a memory that can be efficiently and effectively updated over time, as new facts and external feedback continue t

Human-in-the-Loop Atlas-Based 3D Asset Segmentation for Interactive Content Workflows

ApplicationsDGX agent

arXiv:2606.17824v2 Announce Type: replace-cross Abstract: Segmenting 3D assets into meaningful regions remains challenging, especially when segmentation criteria are application-dependent and require

HyPASE: Hyperbolic Geometry for Parameter-Efficient Speech Emotion Fine-Tuning Framework for Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2608.04351v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) excel at general speech understanding; however, adapting them to fine-grained tasks like Speech Emotion Recognitio

Image Classification Using CNN-QNN Hybrid Model with Optimized Correlated Features

ResearchDGX agent

arXiv:2608.04379v1 Announce Type: cross Abstract: We propose a method to optimize the correlation among convolutional neural network (CNN) features that are used as inputs to quantum neural network (Q

IMFACT: Counterfactual Explanations for Time Series via Intrinsic Mode Function Substitution

ResearchDGX agent

arXiv:2608.04777v1 Announce Type: cross Abstract: Oscillatory signals, such as vibration, carry class-discriminative information in specific frequency bands; perturbing them in raw feature space for c

Improving Auto-Design of Neural PDE Solvers with a Domain-Specific Language

AgentsDGX agent

arXiv:2608.04384v1 Announce Type: new Abstract: Neural PDE solver auto-design is fundamentally a search-space representation problem. In the space of unrestricted Python programs, valid solvers form a

InsightEmb: Learning Action-Intent Embeddings for Agentic Insight Retrieval

Model ReleasesDGX agent

arXiv:2608.04761v1 Announce Type: cross Abstract: Self-improving agents accumulate reusable insights from prior trajectories, making retrieval increasingly important for turning accumulated experience

Interoceptive Attention as Dynamic Homeostatic Prioritization in a Foraging Agent

AgentsDGX agent

arXiv:2608.04232v1 Announce Type: new Abstract: Biological systems must regulate competing needs under limited perceptual bandwidth, where sharpening one estimate costs the capacity to sharpen the oth

Interpretable Fuzzy Inference for UAV Target Tracking Using Bounding-Box Geometry

Local AiDGX agent

arXiv:2608.04121v1 Announce Type: cross Abstract: Vision-based guidance of unmanned aerial vehicles (UAVs) toward unmanned ground vehicles (UGVs) supports cooperative aerial--ground robotics, but reli

Interpreting GFlowNets for Drug Discovery: What probes can and cannot show

SafetyDGX agent

arXiv:2511.19264v2 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) construct molecules through sequential decisions, but their internal policies remain opaque, limiting ado

InvFlowFD: Reference-Free and Background-Set-Free Perceptual Music Quality Metric with Flow Matching Inversion

ResearchDGX agent

arXiv:2608.04142v1 Announce Type: cross Abstract: Existing reference-free methods for evaluating music perceptual quality alleviate the need for paired noisy-clean data, but they still rely on a backg

Is Monitoring Enough? Strategic Agent Selection For Stealthy Attack in Multi-Agent Discussions

AgentsDGX agent

arXiv:2603.21194v2 Announce Type: replace-cross Abstract: Multi-agent discussions have been widely adopted, motivating growing efforts to develop attacks that expose their vulnerabilities. In this wor

iStructTab: Structured Feature Sequencing for Multimodal Learning of Image and Tabular Data

ResearchDGX agent

arXiv:2608.04348v1 Announce Type: cross Abstract: Multimodal learning of images and tabular data is often impaired by ineffective representations, resulting in redundancy, dispersion, and generalizati

Item Response Theory for AI Safety

Model ReleasesDGX agent

arXiv:2608.05086v1 Announce Type: new Abstract: Language models differ in how safely they behave and these differences are measured by safety benchmarks. But aggregated benchmark scores are hard to tr

Joint UAV Flight and Opportunistic Routing under Reinforcement Learning for Delay-Tolerant Networks

SafetyDGX agent

arXiv:2608.04590v1 Announce Type: new Abstract: The growing deployment of delay-tolerant networks (DTNs) has made store-carry-forward (SCF) communication indispensable under sparse connectivity. Howev

LaPrune: Controllable Differentiable Sparsity at Million Scale

Model ReleasesDGX agent

arXiv:2608.04057v1 Announce Type: cross Abstract: Top-k selection determines which components of a sparse model remain active. Hard selection blocks gradients, while continuous relaxations often coupl

Large-Small Model Collaboration for Enhancing Edge-Deployed Small Models

Model ReleasesDGX agent

arXiv:2503.10367v2 Announce Type: replace-cross Abstract: Edge devices host domain-specific small language models (SLMs) with limited resources, while private clouds offer larger LLMs. We propose G-Bo

Leak-Resistant Unlearning: A New Benchmark for Evaluating Multi-Hop Reasoning Consistency and Recovery Robustness

Model ReleasesDGX agent

arXiv:2608.04519v1 Announce Type: new Abstract: Benchmarking machine unlearning methods is critical to understand whether sensitive knowledge is removed from large language models (LLMs) or not. Curre

LiNC: Lightweight Noise Correction via Per-Sample Trust and Gaussian Mixture Modeling

Model ReleasesDGX agent

arXiv:2608.04147v1 Announce Type: cross Abstract: Label noise is common in medical imaging datasets due to factors such as inter-rater variability, annotation errors, and ambiguous cases. This can sev

Lindblad-Inspired Multi-Timescale Reservoir Computing with Separable Rotation and Dissipation

Model ReleasesDGX agent

arXiv:2608.04028v1 Announce Type: cross Abstract: Echo-state networks enable efficient temporal learning by fixing the recurrent dynamics and training only a linear readout. However, conventional rese

Long-term Measurements: Towards a Longitudinal Understanding of Human-AI Interactions

SafetyDGX agent

arXiv:2608.02491v2 Announce Type: replace Abstract: Language models have taken on the role of a very new type of technology, by virtue of their 'human-ness' and rapid integration into users' daily liv

Mamba with Hierarchical Memory: Solving Representation Bottleneck in Long Sequence Modeling

ResearchDGX agent

arXiv:2608.02347v2 Announce Type: replace Abstract: Recurrent linear attention models (RLAs) such as Mamba offer efficient linear-time sequence modeling as an alternative to Transformers, yet their fi

MarsCast: Transfer Learning of AI Weather Foundation Models to Planetary Atmospheres

ResearchDGX agent

arXiv:2608.05054v1 Announce Type: cross Abstract: We investigate the transferability of Earth weather foundation models to planetary atmospheres by adapting the GraphCast graph neural weather forecast

Masked diffusion enables coherent beat tracking

ResearchDGX agent

arXiv:2608.04624v1 Announce Type: cross Abstract: Current neural networks for beat tracking generate invalid outputs, such as consecutive downbeats and erratic tempo changes, even when these are not p

MatrAIx: Simulating the World with 8.3 Billion Persona Agents

Model ReleasesDGX agent

arXiv:2608.04205v1 Announce Type: new Abstract: Human evaluation of AI systems and digital products is costly, slow, and difficult to scale. Offline evaluations are more scalable but often abstract aw

MediRec: Enhancing Chinese Medication Recommendation with Explainable Clinical Reasoning

Model ReleasesDGX agent

arXiv:2510.21084v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong potential for clinical decision support through their advanced language understanding and reaso

MemFly: On-the-Fly Memory Optimization via Information Bottleneck

ResearchDGX agent

arXiv:2602.07885v2 Announce Type: replace Abstract: Long-term memory enables large language model agents to tackle complex tasks through historical interactions. However, existing frameworks encounter

Memorization in Large Language Models in Medicine: Prevalence, Characteristics, and Implications

SafetyDGX agent

arXiv:2509.08604v5 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated significant potential in medicine, with many studies adapting them through continued pre-traini

MERaLiON-GR: Speech Gender Recognition Model for English and SEA Languages

Model ReleasesDGX agent

arXiv:2608.04433v1 Announce Type: cross Abstract: We present MERaLiON-GR, a speech gender recognition system that performs binary classification (female / male) on English and Southeast Asian (SEA) la

MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training

Model ReleasesDGX agent

arXiv:2608.04407v1 Announce Type: cross Abstract: Memory-efficient matrix optimizers such as Sinkhorn gradient descent remove most AdamW optimizer state for dense Transformer matrices, but direct appl

MIDAS: Multi-LLM Iterative Data-Adaptive Summarization

ApplicationsDGX agent

arXiv:2608.04307v1 Announce Type: cross Abstract: Text summarization is deceptively difficult. While condensing information seems straightforward, real-world enterprise summarization of support ticket

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding

Model ReleasesDGX agent

arXiv:2608.04054v1 Announce Type: cross Abstract: Multimodal intent recognition requires understanding not only what textual, acoustic, and visual signals share, but also how they disagree. Such disag

← Previous
1…2728293031…354
Next →