AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accurate Optimization Modeling

DGX agent

arXiv:2608.07040v1 Announce Type: new Abstract: Solving combinatorial optimization problems (COPs) requires not only efficient algorithms but also carefully crafted formulations. While recent works ha

model-releasesarxiv-cs-ai
10 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Policy-Masked Private Experts: Auditable and Reversible Capability Access Control in Sparse MoE Models

DGX agent

arXiv:2608.06690v1 Announce Type: cross Abstract: Most language-model access controls regulate behavior while leaving the same computation available to every request. We study a different systems ques

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Social World Models

DGX agent

arXiv:2509.00559v3 Announce Type: replace Abstract: Humans intuitively navigate social interactions by simulating unspoken dynamics and reasoning about others' perspectives, even with limited informat

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML Comparisons

DGX agent

arXiv:2608.07303v1 Announce Type: new Abstract: Comparisons between AutoML systems at short time budgets -- tens of seconds rather than hours -- are common in tool READMEs and workshop papers, and the

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination

DGX agent

arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores once memorized. extbf{Contamination mitigation

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

Adaptive Arena-based Contestable Argumentative Network-of-Experts for Open-Ended Care Plan Coordination

DGX agent

arXiv:2608.05391v1 Announce Type: new Abstract: Care plan coordination demands synthesizing heterogeneous clinical, functional, and psychosocial information across multiple professional disciplines, w

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Enhancing Social Intelligence in LLMs with Hierarchical Reasoning and Utterance-Level Goal Rewarding

DGX agent

arXiv:2608.05832v1 Announce Type: new Abstract: Large language models (LLMs) excel in structured tasks but struggle with dynamic social interactions, where success requires long-term goal coordination

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

Adversarial Attacks for Good: A Survey of Proactive Protection across the Visual Content Lifecycle

DGX agent

arXiv:2608.04314v1 Announce Type: cross Abstract: Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can add

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding

DGX agent

arXiv:2608.04156v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instru

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

HCRide: Harmonizing Passenger Fairness and Driver Preference for Human-Centered Ride-Hailing

DGX agent

arXiv:2508.04811v2 Announce Type: replace Abstract: Order dispatch systems play a vital role in ride-hailing services, which directly influence operator revenue, driver profit, and passenger experienc

safetyarxiv-cs-lg
6 Aug 2026
Model Releases

K-EXAONE 2.0 Technical Report

DGX agent

arXiv:2608.04505v1 Announce Type: new Abstract: This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward glo

model-releasesarxiv-cs-cl
6 Aug 2026
Safety

NSF-HRPT: Neural Semantic Field meets Hierarchical Risk Perception Tree for Safety-Critical Scenario Assessment

DGX agent

arXiv:2608.04776v1 Announce Type: new Abstract: The ability to accurately assess and anticipate risks in safety-critical scenarios is crucial for autonomous driving systems. While existing research ha

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

The Yokai Learning Environment: Tracking Beliefs Over Space and Time

DGX agent

arXiv:2508.12480v3 Announce Type: replace Abstract: The ability to cooperate with unknown partners is a central challenge in cooperative AI and widely studied in the form of zero-shot coordination (ZS

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Evaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning

DGX agent

arXiv:2608.03028v1 Announce Type: new Abstract: Applying a valid medication-safety rule when its patient-specific conditions are not met can produce an incorrect decision. Existing medical evaluations

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

When and Where to Look: Adaptive Visual Evidence Scheduling for Efficient Long Video Understanding

DGX agent

arXiv:2608.03918v1 Announce Type: cross Abstract: Efficient long-video understanding requires vision--language models (VLMs) to reason over a small number of frames selected as sparse visual evidence.

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

Abduction Without a Body? Representational Grounding and the Abduction Loop for Scientific Hypothesis Generation

DGX agent

arXiv:2608.02505v1 Announce Type: cross Abstract: Can scientific abduction occur without continuous sensorimotor embodiment? Recent arguments in AI and philosophy of science hold that genuine hypothes

model-releasesarxiv-cs-cv
4 Aug 2026
Local Ai

Device-First Feedback: Toward Mobile-Native LLM-Driven Neural Architecture Search

DGX agent

arXiv:2608.00078v1 Announce Type: new Abstract: Deploying convolutional neural networks generated by large language models (LLMs) on real mobile hardware requires more than GPU validation accuracy: IN

local-aiarxiv-cs-cv
4 Aug 2026
Safety

IACM-RL: Intent-Aware Context Management and Reinforcement Learning for Complex Tool Invocation under Dynamic Intent Fluctuations

DGX agent

arXiv:2608.02110v1 Announce Type: new Abstract: Executing long-horizon tool invocations in real-world environments is severely challenged by dynamic user intent noise. Existing methods attempt robustn

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

MixedComplementarityProblems.jl: A Fast, Batched, Open-Source Interior Point Solver for Mixed Complementarity Problems

DGX agent

arXiv:2608.00959v1 Announce Type: cross Abstract: Mixed complementarity problems (MCPs) arise as the first-order optimality conditions of nonlinear programs and noncooperative games, and provide a nat

model-releasesarxiv-cs-ro
4 Aug 2026
Safety

Practical Noise Modeling for SPAD Intensity Imaging

DGX agent

arXiv:2608.00489v1 Announce Type: new Abstract: Single-photon avalanche diode (SPAD) cameras are promising for low-light and high-dynamic-range intensity imaging, but their practical use is limited by

safetyarxiv-cs-cv
4 Aug 2026
Model Releases

SCHEDBench: A Benchmark for Evaluating LLM Constraint Faithfulness in Natural-Language Combinatorial Scheduling

DGX agent

arXiv:2608.00991v1 Announce Type: cross Abstract: This paper introduces SCHEDBench, a natural-language benchmark for evaluating combinatorial scheduling constraint faithfulness under surface-form vari

model-releasesarxiv-cs-cl
4 Aug 2026
Safety

Self-Improving Large Language Models via Progressive Experience Evolution

DGX agent

arXiv:2608.02139v1 Announce Type: new Abstract: Large language models (LLMs) capable of self-improvement require not only effective policy optimization, but also a principled mechanism for transformin

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Toward Robust LLM-Based Judges: Taxonomic Bias Evaluation and Debiasing Optimization

DGX agent

arXiv:2603.08091v2 Announce Type: replace Abstract: Large language model (LLM)-based judges are widely adopted for automated evaluation and reward modeling, yet their judgments are often affected by j

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Benchmarks Are Not Validation: A System-Level View of Financial LLM Applications

DGX agent

arXiv:2607.28840v1 Announce Type: new Abstract: Large language models are increasingly deployed in financial applications that combine retrieval, proprietary data, tool use, orchestration logic, monit

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

Detecting Experiential Intertextuality Across Migration Routes: Beyond Surface Similarity in French Narratives

DGX agent

arXiv:2607.29188v1 Announce Type: new Abstract: Migrants traversing geographically distinct routes such as the Trans-Saharan and Balkan corridors often recount strikingly parallel lived experiences: p

model-releasesarxiv-cs-cl
3 Aug 2026
Safety

Do Medical Foundation Models Generalize on the African Brain?

DGX agent

arXiv:2607.28771v1 Announce Type: new Abstract: Medical foundation models (FMs) are increasingly used for brain MRI analysis. However, their evaluation remains dominated by high-resource datasets, lea

safetyarxiv-cs-cv
3 Aug 2026
Safety

Fragility of Value under Imperfect Alignment

DGX agent

arXiv:2607.28881v1 Announce Type: new Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that these systems are aligned with humanity. A common

safetyarxiv-cs-ai
3 Aug 2026
Safety

Gated Q-learning: Add Off-Policy Bias to Taste

DGX agent

arXiv:2607.28916v1 Announce Type: cross Abstract: Multistep credit assignment is critical for sample-efficient reinforcement learning, yet managing off-policy bias in Q-learning remains a fundamental

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Identifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design

DGX agent

arXiv:2607.28894v1 Announce Type: new Abstract: Computational cognitive modeling seeks to infer latent cognitive mechanisms underlying observed behavior. Bayesian inverse planning provides a principle

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Linear Proposal Operators and Stochastic Search Geometry in SOMA and Differential Evolution

DGX agent

arXiv:2607.29228v1 Announce Type: cross Abstract: Swarm and evolutionary algorithms are usually analyzed as complete procedural systems in which nonlinear selection, replacement, and adaptation obscur

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Looks Right, Works Right: A Project-Level Benchmark for Multi-Screen Mobile App Generation

DGX agent

arXiv:2607.28645v1 Announce Type: cross Abstract: Recent multimodal large language models can convert visual designs directly into executable code, but real mobile products require multiple screenshot

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

Reflection or Re-Generation? Why LLM Revision Fails Where Human Revision Succeeds

DGX agent

arXiv:2607.28908v1 Announce Type: new Abstract: Reflection, the ability to revisit and revise prior reasoning, is central to how humans improve their answers. Large language models (LLMs) are increasi

local-aiarxiv-cs-lg
3 Aug 2026
Model Releases

TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

DGX agent

arXiv:2607.28657v1 Announce Type: new Abstract: Large Language Models (LLMs) often require carefully crafted prompts to unlock their full potential, which can be a barrier for non-expert users. This w

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

To Add Is Machine, To Delete Is Human: Measuring and Mitigating Deletion Avoidance in LLM Code Editing

DGX agent

arXiv:2607.28887v1 Announce Type: cross Abstract: Large language models increasingly write and repair production code, yet evidence is mounting that their test-passing patches leave codebases harder t

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

A Systems Engineering Framework for Vision-Language-Enabled UAV Triage and Disaster Response

DGX agent

arXiv:2607.27597v1 Announce Type: new Abstract: Recent advances in Vision Language Models (VLMs) have created new opportunities for disaster response, where responders must interpret large volumes of

safetyarxiv-cs-ro
31 Jul 2026
Model Releases

AgentMap: Joint Equivalence and Subsumption Discovery for Ontology Matching

DGX agent

arXiv:2607.27130v1 Announce Type: new Abstract: Ontology matching (OM) has traditionally been formulated as either equivalence discovery or subsumption matching. The existing OM systems identify only

model-releasesarxiv-cs-ai
31 Jul 2026
Safety

Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak Defenses

DGX agent

arXiv:2607.26639v1 Announce Type: cross Abstract: A self-check defense asks the target model to assess a request before answering it; SAGE, the strongest published instance, reports an average 99% def

safetyarxiv-cs-ai
31 Jul 2026
Model Releases

Bridging AI and Energy Forecasting: An Autonomous Workflow with Customized Toolkit

DGX agent

arXiv:2307.07191v3 Announce Type: replace Abstract: Energy forecasting is crucial for the power grid, but fundamentally different from general time series analysis: it highly relies on covariates like

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Can LVLMs Uncover the Truth Behind Visual Illusions? An Analysis of Perceptual and Reasoning Capabilities

DGX agent

arXiv:2607.27747v1 Announce Type: new Abstract: Large Vision Language Models have integrated reasoning capabilities, elevating cognitive performance to new levels. However, existing evaluations either

model-releasesarxiv-cs-cl
31 Jul 2026
Local Ai

Can Vision-Language Models Reason about AI Edits in Images?

DGX agent

arXiv:2607.28464v1 Announce Type: new Abstract: Detection and localization of AI-tampered images are critical for trustworthy AI, yet modern generative models have made such manipulations increasingly

local-aiarxiv-cs-cv
31 Jul 2026
Safety

Digital Harf: A Clinically Integrated Multimodal AI System for Pervasive Arabic Speech and Language Therapy

DGX agent

arXiv:2607.27212v1 Announce Type: cross Abstract: Children with Autism Spectrum Disorder in Arabic-speaking countries face compounded barriers to effective speech and language therapy: a shortage of q

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

DoTime: A Synthetic Benchmark Generator for Interventional and Counterfactual Time Series

DGX agent

arXiv:2607.27263v1 Announce Type: new Abstract: Most benchmarks for causal inference over time series are observational, small, or domain-specific, leaving interventional and counterfactual estimation

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adaptation in Coding Assistants

DGX agent

arXiv:2607.26611v1 Announce Type: new Abstract: AI-assisted coding increasingly translates informal user intent into executable software, yet coding requests often contain ambiguities that recur in us

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

From Single- to Cross-Document: Benchmarking Multi-Granularity Event Analysis of Large Language Models

DGX agent

arXiv:2607.27654v1 Announce Type: new Abstract: Event analysis is an essential and fundamental direction of information extraction, involving various event-centric tasks at different granularity of do

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

LLMs struggle to simulate human belief updates in controlled environments

DGX agent

arXiv:2607.28347v1 Announce Type: new Abstract: LLMs are increasingly deployed as proxies for human study participants in social science experiments, yet the fidelity of this practice has rarely been

model-releasesarxiv-cs-cl
31 Jul 2026
Hardware

ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding

DGX agent

arXiv:2607.28312v1 Announce Type: new Abstract: Streaming video understanding requires models to continuously retain useful visual evidence before future questions are known. Existing approaches prima

hardwarearxiv-cs-cv
31 Jul 2026
Model Releases

RepBench: Compiling Benchmarks into Capability Representations for Large Language Models

DGX agent

arXiv:2607.28008v1 Announce Type: new Abstract: Representation engineering reads and steers capability directions in large language models, yet methods are typically evaluated on paper-specific synthe

model-releasesarxiv-cs-cl
31 Jul 2026
Tutorials

RLPF: Reinforcement Learning from Performance Feedback for Code Generation

DGX agent

arXiv:2607.27271v1 Announce Type: new Abstract: Code models are increasingly trained with execution feedback, but most training signals still stop at correctness. This leaves an important gap for syst

tutorialsarxiv-cs-lg
31 Jul 2026
← Previous
1…186187188189190…233
Next →