AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
83,193 results
11 Aug 2026

PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling

Model ReleasesDGX agent

arXiv:2608.08700v1 Announce Type: new Abstract: Reliable evaluation of tool routing is critical as Large Language Models increasingly operate as autonomous agents. Current benchmarks face three struct

PolicyKG: An Agentic LLM Pipeline for Translating Institutional Policies into SHACL Knowledge Graphs

Model ReleasesDGX agent

arXiv:2608.09028v1 Announce Type: new Abstract: Institutional policies stay in natural language while the systems that check compliance demand machine-readable constraints. Bridging that gap is still

PolypSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering

ResearchDGX agent

arXiv:2603.07066v2 Announce Type: replace-cross Abstract: Generative diffusion models are increasingly used for medical imaging data augmentation, but text prompting cannot produce causal training dat

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Population-Level Generative Modeling for Ranking Data

Model ReleasesDGX agent

arXiv:2608.08422v1 Announce Type: cross Abstract: Ranking data arise in scientific and machine learning applications, including recommendation systems, information retrieval, voting, marketing, and AI

Population-Scalable Multi-Agent World Modeling

AgentsDGX agent

arXiv:2608.08600v1 Announce Type: cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environment

PosBridge: Multi-View Positional Embedding Transplant for Identity-Aware Image Editing

Local AiDGX agent

arXiv:2508.17302v2 Announce Type: replace Abstract: Localized subject-driven image editing aims to seamlessly integrate user-specified objects into target scenes. As generative models continue to scal

Position Bias in Ordinal Classification: A Systematic Evaluation

SafetyDGX agent

arXiv:2608.08869v1 Announce Type: new Abstract: Large language models are increasingly used for ordinal classification, yet semantically equivalent changes to prompt organization can alter their predi

Position: Certifiable State Integrity Should Be Built from Local Validity, Not Global Scale

Local AiDGX agent

arXiv:2601.21249v2 Announce Type: replace Abstract: Breakthroughs in language and vision have motivated increasingly general foundation models for time series and physical dynamics, where evidence is

Positioning Generative Artificial Intelligence in STEM Assessment: When to Require, Scaffold, or Restrict Its Use

ResearchDGX agent

arXiv:2608.07475v1 Announce Type: cross Abstract: Generative Artificial Intelligence (GenAI) presents a governance challenge for STEM assessment. Unrestricted access can enable task outsourcing that u

PQC in Plaintext: Google Cloud’s post-quantum cryptography roadmap

SafetyDGX agent

Securing infrastructure and services against a future cryptographically-relevant quantum computer has been a goal for Google for a decade, and we’ve dedicated ourselves to help developers by advancing

PragMatch: Separating Pragmatic Incongruity from Cross-Modal Mismatch in Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2608.09772v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated strong performance on multimodal benchmarks, yet it remains unclear whether they genuinely reason

Pragmatic Attack Surface: Vulnerabilities of Implicit Context in Large Language Models

SafetyDGX agent

arXiv:2608.09551v1 Announce Type: new Abstract: In the era of large language models (LLMs), attackers often manipulate natural language to elicit unsafe or harmful outputs, creating a new natural lang

PragyaDoc: A Universal Document Intelligence Framework for Multilingual Medical Document Understanding in Low-Resource Settings

ResearchDGX agent

arXiv:2608.07478v1 Announce Type: cross Abstract: India's 22 official languages create a critical accessibility barrier: the majority of medical documentation exists exclusively in English, yet the pa

Predict to Skip: Linear Multistep Feature Forecasting for Efficient Diffusion Transformers

ResearchDGX agent

arXiv:2602.18093v2 Announce Type: replace Abstract: Diffusion Transformers (DiT) have emerged as a widely adopted backbone for high-fidelity image and video generation, yet their iterative denoising p

Predicting blood clot growth from sparse post-onset measurements with latent neural differential equations

Model ReleasesDGX agent

arXiv:2608.08165v1 Announce Type: new Abstract: Computational models of blood clotting improve understanding of thrombus formation, but their clinical application remains limited because many model in

Predictive Failure Detection in Network Hardware Using Thermal Imaging and Deep Learning with Sensor Fusion

ResearchDGX agent

arXiv:2608.07582v1 Announce Type: new Abstract: Unplanned network hardware malfunctions can interrupt services and result in expensive downtime in data centers. A deep learning-based predictive mainte

Predictive safety filter enhanced curriculum learning control for efficient vehicle dynamics controller

Model ReleasesDGX agent

arXiv:2608.09653v1 Announce Type: cross Abstract: Recent advances in learning-based control have enabled impressive achievements in solving complex control problems in various domains. However, since

Preference Redirection via Attention Concentration: An Attack on Computer Use Agents

AgentsDGX agent

arXiv:2604.08005v2 Announce Type: replace Abstract: Advancements in multimodal foundation models have enabled the development of Computer Use Agents (CUAs) capable of autonomously interacting with GUI

PreGress: Ranking-Native Pre-training and Prompting for Graph Node Ranking

ApplicationsDGX agent

arXiv:2608.09016v1 Announce Type: cross Abstract: Node ranking is a fundamental problem in graph information retrieval, measuring the relative importance of nodes and supporting a wide range of applic

Preserve More Details: Mitigating Content Drift in Real-World Image Super-Resolution

ApplicationsDGX agent

arXiv:2608.09373v1 Announce Type: new Abstract: Real-world image super-resolution (Real-ISR) aims to reconstruct high-quality (HQ) images from low-quality (LQ) inputs subject to diverse real-world deg

Preserving Item Semantics for Free: Rethinking Token Initialization in LLM-Based Generative Recommendation

Model ReleasesDGX agent

arXiv:2608.07816v1 Announce Type: cross Abstract: Recent advances in generative recommendation (GR) leverage large language models (LLMs) as recommender backbones, enabling LLMs to directly generate r

PressureMesh: 3D Human Mesh Estimation from Multi-Device Pressure Images

ResearchDGX agent

arXiv:2608.09550v1 Announce Type: new Abstract: Human pose monitoring is crucial in fields such as rehabilitation assessment and human-computer interaction. Due to its privacy-preserving nature, press

Preview-Based Relative-Motion Control of an Insertion Tool for Neural-Thread Placement in Pulsating Tissue

ResearchDGX agent

arXiv:2608.08860v1 Announce Type: cross Abstract: Robotic neural-thread placement requires regulating the insertion-tool tip relative to tissue that moves with cardiac and respiratory pulsation. This

PRISM: A Predictive Protocol for Permutation Optimization via Landscape Diagnostics

ResearchDGX agent

arXiv:2608.08344v1 Announce Type: cross Abstract: Permutation optimization arises whenever the components of a system are fixed but their ordering affects performance. We introduce PRISM, a predictive

PRISM-Delta: Differential Subspace Steering for Prompt Highlighting in Large Language Models

ResearchDGX agent

arXiv:2603.10705v2 Announce Type: replace Abstract: Prompt highlighting steers a large language model to prioritize user-specified text spans during generation. A key challenge of existing Key-editing

Privacy-Preserving Data Drift Detection and Recovery for Large-Scale LLM Applications via Proxy Representations

SafetyDGX agent

arXiv:2608.08245v1 Announce Type: cross Abstract: LLM applications deployed at scale face a fundamental challenge: privacy constraints prevent direct inspection of user interactions, making it difficu

Private Anytime Selective-Risk Certification for Federated Retrieval-Augmented Generation: Guarantees and Empirical Limits

SafetyDGX agent

arXiv:2608.07913v1 Announce Type: cross Abstract: Selective-risk certificates promise that accepted outputs meet a declared error target. We develop Fed-SRC, a score-agnostic certificate for federated

Private Etymology: Designing Relational Reuse of Shared Symbols in Long-Term Human-AI Interaction

Local AiDGX agent

arXiv:2608.08443v1 Announce Type: cross Abstract: Previous studies have shown that people can develop shared symbols, partner-specific expressions, personal idioms, inside jokes, and other parts of a

Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Policy Self-Distillation

SafetyDGX agent

arXiv:2608.09263v1 Announce Type: new Abstract: Outcome verifiers score completed reasoning traces but do not assign credit to intermediate tokens. Privileged self-distillation attempts to fill this g

Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation

SafetyDGX agent

arXiv:2608.09228v1 Announce Type: cross Abstract: On-Policy Self-Distillation (OPSD) is commonly interpreted as the transfer of privileged information: a teacher observes the verified solution to the

Probabilistic Circuits for Knowledge Graph Completion with Reduced Rule Sets

Model ReleasesDGX agent

arXiv:2508.06706v2 Announce Type: replace Abstract: Rule-based methods for knowledge graph completion provide explainable results, but often require tens of thousands of rules to achieve competitive p

ProbSPARQL: Querying Knowledge Graphs with Multi-dimensional, Uncertain Numeric Data

ResearchDGX agent

arXiv:2607.18262v2 Announce Type: replace Abstract: The SFB 1574 Circular Factory is building a shared knowledge graph infrastructure for integrating data about returned products. A central challenge

Progressive Learned Image Compression for Machine Perception

ApplicationsDGX agent

arXiv:2512.20070v2 Announce Type: replace Abstract: Recent advances in learned image codecs have extended from human perception toward machine perception However, progressive image compression with fi

Projection-Retraction MPPI: Exact Constraint-Manifold Control for Manipulators

ResearchDGX agent

arXiv:2608.07573v1 Announce Type: new Abstract: Model Predictive Path Integral (MPPI) control is widely used in manipulation for its gradient-free, parallel handling of non-convex costs. Manipulation

Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States

ResearchDGX agent

arXiv:2608.08024v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent and useful responses but remain prone to hallucinations. We introduce Prompt Embedding Probes (PEP),

Prompt engineering does not universally improve Large Language Model performance across clinical decision-making tasks

Model ReleasesDGX agent

arXiv:2512.22966v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated promise in medical knowledge assessments, yet their practical utility in real-world clinical decision

ProPINN: Demystifying Propagation Failures in Physics-Informed Neural Networks

ResearchDGX agent

arXiv:2502.00803v3 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have earned high expectations in solving partial differential equations (PDEs), but their optimization usua

PROSLEX: A Novel Dataset for Expert-Annotated Legal Statute Prediction for Indian Judiciary

Model ReleasesDGX agent

arXiv:2608.08830v1 Announce Type: new Abstract: Legal Statute Prediction (LSP) involves automatically identifying relevant legal statutes given factual descriptions in legal documents, typically frame

Prospects of Finding a ML Engineering Job [D]

ResearchDGX agent

Hello all, I am wondering if a transition from a Ph.D. in electrical engineering (Quantum optics/photonics) to a job in ML is a reasonable aspiration. Personally, I have extensive software development

Protecting patient privacy in clinical foundation models: Technical and legal perspectives

ApplicationsDGX agent

arXiv:2608.07705v1 Announce Type: new Abstract: Clinical foundation models trained on large-scale patient data are increasingly used for decision support, screening, and public health. As deployment e

Proxy OPD: On-Policy Distillation with Transferable Relative Proxy Update

SafetyDGX agent

arXiv:2607.11505v2 Announce Type: replace-cross Abstract: Post-training for large language models typically couples policy exploration with model optimization, hindering the reuse of high-reward behav

Psychological methods really work. Has anyone tried to encouraging and instilling confidence to GPT?

Model ReleasesDGX agent

Oddly enough, encouragement actually influences the performance of not only Claude but also GPT and other AIs. While Claude was working on a complex problem related to the Riemann Hypothesis, the Anth

Quality-Diversity Stress Tests for Process Reward Models:What Archive Coverage Can and Cannot Certify

ResearchDGX agent

arXiv:2608.08008v1 Announce Type: new Abstract: Process reward models (PRMs) score intermediate reasoning steps and are widely used for search, ranking, and training, but optimization can exploit thes

Quantifying Membership Disclosure Risk for Tabular Synthetic Data Using Kernel Density Estimators

ApplicationsDGX agent

arXiv:2603.10937v2 Announce Type: replace Abstract: The use of synthetic data has become increasingly popular as a privacy-preserving alternative to sharing real datasets, especially in sensitive doma

Quantization Degradation in Large Language Models: A Signal-Noise Perspective

ResearchDGX agent

arXiv:2608.08188v1 Announce Type: new Abstract: Post-training quantization reduces the deployment cost of large language models, yet how severely a quantized model degrades is not determined by bit-wi

QuantumMind: Constraint-Grounded Agentic Reasoning for Speedup Analysis in Quantum Computing

AgentsDGX agent

arXiv:2608.07743v1 Announce Type: new Abstract: Identifying a meaningful quantum speedup requires more than matching a classical problem to a familiar quantum primitive: the claim must preserve the ta

QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture

Model ReleasesDGX agent

arXiv:2510.22087v3 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from

Query-Only Backdoor Attacks on Self-Evolving Skills via Trajectory Poisoning

AgentsDGX agent

arXiv:2608.08303v1 Announce Type: new Abstract: Agentic skills improve large language model (LLM) agents by encoding reusable procedures for complex tasks. However, manually authored skills often adap

Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis

Model ReleasesDGX agent

arXiv:2509.21629v4 Announce Type: replace-cross Abstract: Program verification relies on loop invariants, yet automatically discovering strong invariants remains a long-standing challenge. We investig

R3S: Refining and Recovering Reinforcement Signals for Multilingual Understanding and Reasoning

ResearchDGX agent

arXiv:2602.05940v2 Announce Type: replace Abstract: Large reasoning models often default to English reasoning when processing non-English questions, yet their performance drops substantially when reas

RA-FinBERT: Rule-aware LoRA adaptation for low-resource financial sentiment classification

Model ReleasesDGX agent

arXiv:2608.09834v1 Announce Type: new Abstract: Financial sentiment analysis converts unstructured financial news into quantitative signals that can support market analysis and decision-making. Existi

RAG-3DSG: Enhancing 3D Scene Graphs with Re-Shot Guided Retrieval-Augmented Generation

ApplicationsDGX agent

arXiv:2601.10168v3 Announce Type: replace-cross Abstract: Open-vocabulary 3D Scene Graph (3DSG) can enhance various downstream tasks in robotics by leveraging structured semantic representations, yet

RAG-Audio: Retrieval-Augmented Generation for Faithful Brain-to-Audio Reconstruction

ResearchDGX agent

arXiv:2608.09331v1 Announce Type: cross Abstract: Brain-to-audio reconstruction is limited by prior domination: when a pretrained generator is conditioned on a weak neural signal, it produces realisti

RAG-Based Auto-Configuration for Industrial Fieldbus Devices

Model ReleasesDGX agent

arXiv:2608.08618v1 Announce Type: cross Abstract: Industrial device commissioning requires engineers to manually extract hundreds of protocol-specific parameters from heterogeneous PDF manuals and tra

RAGMesh with FaME-G2E: Long-Form Text-Driven 3D Face Generation and Editing

Model ReleasesDGX agent

arXiv:2608.09186v1 Announce Type: new Abstract: Text-driven 3D face generation and editing remains challenging due to the difficulty of translating long-form descriptions into fine-grained facial geom

RangeFactory: Scalable Construction of Multi-Hop Cyber Ranges

AgentsDGX agent

arXiv:2608.09526v1 Announce Type: cross Abstract: Real-world cyberattacks often require sustained progress across multiple hosts and network segments, making multi-hop cyber ranges essential infrastru

RankGuide: Tensor-Rank-Guided Routing and Steering for Efficient Reasoning

ResearchDGX agent

arXiv:2604.16694v2 Announce Type: replace Abstract: Large reasoning models (LRMs) enhance problem-solving capabilities by generating explicit multi-step chains of thought (CoT) reasoning; however, the

RAVEN-Eval: Rubric-Guided Automatic Evaluation for AI Video Generation Models Based on LMM Preference Judgement

ResearchDGX agent

arXiv:2608.09111v1 Announce Type: new Abstract: AI video generation has advanced rapidly and entered widespread commercial use. As a result, quality differences among videos produced by state-of-the-a

RAVEN: Frozen Random Graph Reservoirs with Physics-Informed Interaction Fingerprints for Protein-Ligand Binding Affinity Prediction

ResearchDGX agent

arXiv:2608.09099v1 Announce Type: new Abstract: Quantitative estimation of protein-ligand binding affinity from three-dimensional complex structures is a fundamental task in structure-based computatio

RayLift: Lifting Complementary Ray-Wise Evidence with 3D Geometry Priors for Semantic Scene Completion

Local AiDGX agent

arXiv:2608.08476v1 Announce Type: new Abstract: Camera-based 3D semantic scene completion (SSC) provides comprehensive scene understanding for autonomous driving and robotics. However, existing method

← Previous
1…2728293031…1387
Next →