AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
23,195 results
Model Releases

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

DGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

model-releasesarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

DGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

OptunaHub: A Platform for Black-Box Optimization

DGX agent

arXiv:2510.02798v2 Announce Type: replace Abstract: Black-box optimization (BBO) underpins advances in domains such as AutoML and Materials Informatics, yet implementations of algorithms and benchmark

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PBSBench: A Multi-Level Vision-Language Framework and Benchmark for Hematopathology Whole Slide Image Interpretation

DGX agent

arXiv:2604.17570v1 Announce Type: new Abstract: Peripheral Blood Smear (PBS) is a critical microscopic examination in hematopathology that yields whole-slide imaging (WSI). Unlike solid tissue patholo

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

PFDelta: A Benchmark Dataset for Power Flow under Load, Generation, and Topology Variations

DGX agent

arXiv:2510.22048v3 Announce Type: replace Abstract: Power flow (PF) calculations are the backbone of real-time grid operations, across workflows such as contingency analysis (where repeated PF evaluat

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks

DGX agent

arXiv:2602.06663v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have shown impressive capabilities in generating natural images and supporting multimodal reasoning. However, their

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Plasticity Loss in Deep Reinforcement Learning: A Survey

DGX agent

arXiv:2411.04832v3 Announce Type: replace-cross Abstract: Plasticity refers to a network's ability to adapt to changing data distributions, which is crucial for the successful training of deep reinfor

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention

DGX agent

arXiv:2506.13674v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods have become crucial for rapidly adapting large language models (LLMs) to downstream tasks. Prefix-Tun

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Pulse Shape Discrimination Algorithms: Survey and Benchmark

DGX agent

arXiv:2508.02750v2 Announce Type: replace Abstract: This review presents a comprehensive survey and benchmark of pulse shape discrimination (PSD) algorithms for radiation detection, classifying nearly

model-releasesarxiv-cs-lg
21 Apr 2026
Tutorials

PyEPO: A PyTorch-based End-to-End Predict-then-Optimize Library for Linear and Integer Programming

DGX agent

arXiv:2206.14234v3 Announce Type: replace-cross Abstract: In deterministic optimization, it is typically assumed that all problem parameters are fixed and known. In practice, however, some parameters

tutorialsarxiv-cs-lg
21 Apr 2026
Safety

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

DGX agent

arXiv:2506.07826v2 Announce Type: replace Abstract: Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Tradi

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval

DGX agent

arXiv:2510.08252v2 Announce Type: replace-cross Abstract: In this paper, we introduce ReasonEmbed, a novel text embedding model developed for reasoning-intensive document retrieval. Our work includes

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Rethinking Meeting Effectiveness: A Benchmark and Framework for Temporal Fine-grained Automatic Meeting Effectiveness Evaluation

DGX agent

arXiv:2604.17260v1 Announce Type: new Abstract: Evaluating meeting effectiveness is crucial for improving organizational productivity. Current approaches rely on post-hoc surveys that yield a single c

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ReTrack: Evidence-Driven Dual-Stream Directional Anchor Calibration Network for Composed Video Retrieval

DGX agent

arXiv:2604.17898v1 Announce Type: new Abstract: With the rapid growth of video data, Composed Video Retrieval (CVR) has emerged as a novel paradigm in video retrieval and is receiving increasing atten

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

R&F-Inventory: A Large-Scale Dataset for Monotonic Inventory Estimation in Reach and Frequency Advertising

DGX agent

arXiv:2604.16821v1 Announce Type: new Abstract: Reach and Frequency (R&F) contract advertising is an important form of widely used brand advertising. Unlike performance advertising, R&F contracts emph

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Same Claim, Different Judgment: Benchmarking Scenario-Induced Bias in Multilingual Financial Misinformation Detection

DGX agent

arXiv:2601.05403v2 Announce Type: replace Abstract: Large language models (LLMs) have been widely applied across various domains of finance. Since their training data are largely derived from human-au

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

SciDraw-6K: A Multilingual Scientific Illustration Dataset Generated by Google Gemini

DGX agent

arXiv:2604.17206v1 Announce Type: new Abstract: We present SciDraw-6K, a curated dataset of 6,291 scientific illustrations synthesized by Google Gemini image-generation models, each paired with prompt

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

StealthGraph: Exposing Domain-Specific Risks in LLMs through Knowledge-Graph-Guided Harmful Prompt Generation

DGX agent

arXiv:2601.04740v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in specialized domains such as finance and healthcare, where they introduce unique safety risk

safetyarxiv-cs-cl
21 Apr 2026
Safety

Structure-Aware Diversity Pursuit as an AI Safety Strategy against Homogenization

DGX agent

arXiv:2601.06116v2 Announce Type: replace-cross Abstract: Generative AI models reproduce the biases in the training data and can further amplify them through mode collapse. We refer to the resulting h

safetyarxiv-cs-cl
21 Apr 2026
Safety

The Impact of Off-Policy Training Data on Probe Generalisation

DGX agent

arXiv:2511.17408v4 Announce Type: replace-cross Abstract: Probing has emerged as a promising method for monitoring large language models (LLMs), enabling cheap inference-time detection of concerning b

safetyarxiv-cs-lg
21 Apr 2026
Agents

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey

DGX agent

arXiv:2509.02547v5 Announce Type: replace-cross Abstract: The emergence of agentic reinforcement learning (Agentic RL) marks a paradigm shift from conventional reinforcement learning applied to large

agentsarxiv-cs-cl
21 Apr 2026
Safety

Toward Reusability of AI Models Using Dynamic Updates of AI Documentation

DGX agent

arXiv:2604.17626v1 Announce Type: cross Abstract: This work addresses the challenge of disseminating reusable artificial intelligence (AI) models accompanied by AI documentation (a.k.a., AI model card

safetyarxiv-cs-cl
21 Apr 2026
Tutorials

Towards a Foundation-Model Paradigm for Aerodynamic Prediction in Three-dimensional Design

DGX agent

arXiv:2604.18062v1 Announce Type: new Abstract: Accurate machine-learning models for aerodynamic prediction are essential for accelerating shape optimization, yet remain challenging to develop for com

tutorialsarxiv-cs-lg
21 Apr 2026
Model Releases

Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training

DGX agent

arXiv:2603.23885v3 Announce Type: replace Abstract: Document parsing has recently advanced with multimodal large language models (MLLMs) that directly map document images to structured outputs. Tradit

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

TransXion: A High-Fidelity Graph Benchmark for Realistic Anti-Money Laundering

DGX agent

arXiv:2604.17420v1 Announce Type: new Abstract: Money laundering poses severe risks to global financial systems, driving the widespread adoption of machine learning for transaction monitoring. However

model-releasesarxiv-cs-lg
21 Apr 2026
Applications

UniDomain: Pretraining a Unified PDDL Domain from Real-World Demonstrations for Generalizable Robot Task Planning

DGX agent

arXiv:2507.21545v3 Announce Type: replace Abstract: Robotic task planning in real-world environments requires reasoning over implicit constraints from language and vision. While LLMs and VLMs offer st

applicationsarxiv-cs-ro
21 Apr 2026
Model Releases

VIDS: A Verified Imaging Dataset Standard for Medical AI

DGX agent

arXiv:2604.17525v1 Announce Type: cross Abstract: Medical imaging AI development is fundamentally dependent on annotated datasets, yet no existing standard provides machine-enforceable validation acro

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

WeatherArchive-Bench: Benchmarking Retrieval-Augmented Reasoning for Historical Weather Archives

DGX agent

arXiv:2510.05336v2 Announce Type: replace Abstract: Historical archives on weather events are collections of enduring primary source records that offer rich, untapped narratives of how societies have

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data Generators

DGX agent

arXiv:2602.19946v4 Announce Type: replace Abstract: Recent text-to-image (T2I) diffusion models produce visually stunning images and demonstrate excellent prompt following. But do they perform well as

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

ZoFia: Zero-Shot Fake News Detection with Entity-Guided Retrieval and Multi-LLM Interaction

DGX agent

arXiv:2511.01188v2 Announce Type: replace Abstract: The rapid spread of fake news threatens social stability and public trust, highlighting the urgent need for its effective detection. Although large

safetyarxiv-cs-cl
21 Apr 2026
Tutorials

A Comparative Study on the Impact of Traditional Learning and Interactive Learning on Students' Academic Performance and Emotional Well-Being

DGX agent

arXiv:2604.15335v1 Announce Type: cross Abstract: The growing adoption of interactive learning tools in higher education offers new opportunities to enhance student performance and well-being. This st

tutorialsarxiv-cs-ai
20 Apr 2026
Tutorials

A methodology to rank importance of frequencies and channels in electromyography data with Decision Tree classifiers

DGX agent

arXiv:2604.15353v1 Announce Type: cross Abstract: This study presents a methodology for identifying the most informative frequencies and channels in electromyography (EMG) data to evaluate muscle reco

tutorialsarxiv-cs-lg
20 Apr 2026
Applications

AI-assisted Protocol Information Extraction For Improved Accuracy and Efficiency in Clinical Trial Workflows

DGX agent

arXiv:2602.00052v2 Announce Type: replace-cross Abstract: Increasing clinical trial protocol complexity, amendments, and challenges around knowledge management create significant burden for trial team

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

AISysRev -- LLM-based Tool for Title-abstract Screening

DGX agent

arXiv:2510.06708v3 Announce Type: replace-cross Abstract: Conducting systematic reviews is laborious. In the screening or study selection phase, the number of papers can be overwhelming. Recent resear

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Applied Explainability for Large Language Models: A Comparative Study

DGX agent

arXiv:2604.15371v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult t

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

Beyond MCQ: An Open-Ended Arabic Cultural QA Benchmark with Dialect Variants

DGX agent

arXiv:2510.24328v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used to answer everyday questions, yet their performance on culturally grounded and dialectal co

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

Bilevel Optimization of Agent Skills via Monte Carlo Tree Search

DGX agent

arXiv:2604.15709v1 Announce Type: new Abstract: Agent exttt{skills} are structured collections of instructions, tools, and supporting resources that help large language model (LLM) agents perform part

agentsarxiv-cs-ai
20 Apr 2026
Safety

Cognitive Agency Surrender: Defending Epistemic Sovereignty via Scaffolded AI Friction

DGX agent

arXiv:2603.21735v2 Announce Type: replace-cross Abstract: The proliferation of Generative Artificial Intelligence has transformed benign cognitive offloading into a systemic risk of cognitive agency s

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

CTSCAN: Evaluation Leakage in Chest CT Segmentation and a Reproducible Patient-Disjoint Benchmark

DGX agent

arXiv:2604.15561v1 Announce Type: cross Abstract: Reported chest CT segmentation performance can be strongly inflated when train and test partitions mix slices from the same study. We present CTSCAN,

model-releasesarxiv-cs-cv
20 Apr 2026
Applications

Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning

DGX agent

arXiv:2604.16029v1 Announce Type: new Abstract: Parallel reasoning enhances Large Reasoning Models (LRMs) but incurs prohibitive costs due to futile paths caused by early errors. To mitigate this, pat

applicationsarxiv-cs-cl
20 Apr 2026
Model Releases

DASB -- Discrete Audio and Speech Benchmark

DGX agent

arXiv:2406.14294v3 Announce Type: replace-cross Abstract: Discrete audio tokens have recently gained considerable attention for their potential to bridge audio and language processing, enabling multim

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

DataCenterGym: A Physics-Grounded Simulator for Multi-Objective Data Center Scheduling

DGX agent

arXiv:2604.15594v1 Announce Type: cross Abstract: Modern datacenters schedule heterogeneous workloads across geo-distributed sites with diverse compute capacities, electricity prices, and thermal cond

local-aiarxiv-cs-ai
20 Apr 2026
Local Ai

DINOv3 Beats Specialized Detectors: A Simple Foundation Model Baseline for Image Forensics

DGX agent

arXiv:2604.16083v1 Announce Type: new Abstract: With the rapid advancement of deep generative models, realistic fake images have become increasingly accessible, yet existing localization methods rely

local-aiarxiv-cs-cv
20 Apr 2026
Agents

Discover and Prove: An Open-source Agentic Framework for Hard Mode Automated Theorem Proving in Lean 4

DGX agent

arXiv:2604.15839v1 Announce Type: new Abstract: Most ATP benchmarks embed the final answer within the formal statement -- a convention we call 'Easy Mode' -- a design that simplifies the task relative

agentsarxiv-cs-ai
20 Apr 2026
Safety

DyTact: Capturing Dynamic Contacts in Hand-Object Manipulation

DGX agent

arXiv:2506.03103v2 Announce Type: replace Abstract: Reconstructing dynamic hand-object contacts is essential for realistic manipulation in AI character animation, XR, and robotics, yet it remains chal

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

'Excuse me, may I say something...' CoLabScience, A Proactive AI Assistant for Biomedical Discovery and LLM-Expert Collaborations

DGX agent

arXiv:2604.15588v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into scientific workflows presents exciting opportunities to accelerate biomedical discovery. However,

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation

DGX agent

arXiv:2507.18756v2 Announce Type: replace Abstract: Multi-Armed Bandit (MAB) algorithms are widely used in recommender systems that require continuous, incremental learning. A core aspect of MABs is t

safetyarxiv-cs-lg
20 Apr 2026
Model Releases

Exploring the Capability Boundaries of LLMs in Mastering of Chinese Chouxiang Language

DGX agent

arXiv:2604.15841v1 Announce Type: new Abstract: While large language models (LLMs) have achieved remarkable success in general language tasks, their performance on Chouxiang Language, a representative

model-releasesarxiv-cs-cl
20 Apr 2026
← Previous
1…475476477478479…484
Next →