AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “tools”

GridTimelineEvolution
5,141 results
Safety

PrinciplismQA: A Philosophy-Grounded Approach to Assessing LLM-Human Clinical Medical Ethics Alignment

DGX agent

arXiv:2508.05132v2 Announce Type: replace Abstract: As medical LLMs transition to clinical deployment, assessing their ethical reasoning capability becomes critical. While achieving high accuracy on k

safetyarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models

DGX agent

arXiv:2601.15220v2 Announce Type: replace Abstract: We identify a novel phenomenon in language models: benign fine-tuning of frontier models can lead to privacy collapse. We find that diverse, subtle

safetyarxiv-cs-cl
21 Apr 2026
Tutorials

PyEPO: A PyTorch-based End-to-End Predict-then-Optimize Library for Linear and Integer Programming

DGX agent

arXiv:2206.14234v3 Announce Type: replace-cross Abstract: In deterministic optimization, it is typically assumed that all problem parameters are fixed and known. In practice, however, some parameters

tutorialsarxiv-cs-lg
21 Apr 2026
Local Ai

Q-DeepSight: Incentivizing Thinking with Images for Image Quality Assessment and Refinement

DGX agent

arXiv:2604.16858v1 Announce Type: new Abstract: Image Quality Assessment (IQA) models are increasingly deployed as perceptual critics to guide generative models and image restoration. This role demand

local-aiarxiv-cs-cv
21 Apr 2026
Research

QuickScope: Certifying Hard Questions in Dynamic LLM Benchmarks

DGX agent

arXiv:2604.17842v1 Announce Type: new Abstract: LLM benchmarks are increasingly dynamic: instead of containing a fixed set of questions, they define templates and parameters that can generate an effec

researcharxiv-cs-cl
21 Apr 2026
Model Releases

ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

DGX agent

arXiv:2603.05863v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have revolutionized code generation, standard ``System 1'' approaches that generate solutions in a single forward

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

SQL Query Engine: A Self-Healing LLM Pipeline for Natural Language to PostgreSQL Translation

DGX agent

arXiv:2604.16511v1 Announce Type: cross Abstract: We present SQL Query Engine, an open-source, self-hosted service that translates natural language questions into validated PostgreSQL queries through

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

StepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning

DGX agent

arXiv:2604.18401v1 Announce Type: new Abstract: General agents have given rise to phenomenal applications such as OpenClaw and Claude Code. As these agent systems (a.k.a. Harnesses) strive for bolder

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

SynthFix: Adaptive Neuro-Symbolic Code Vulnerability Repair

DGX agent

arXiv:2604.17184v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for automated code repair but often struggle with the complex semantic and structural correctness required.

tutorialsarxiv-cs-lg
21 Apr 2026
Research

The GDN-CC Dataset: Automatic Corpus Clarification for AI-enhanced Democratic Citizen Consultations

DGX agent

arXiv:2601.14944v3 Announce Type: replace Abstract: LLMs are ubiquitous in modern NLP, and while their applicability extends to texts produced for democratic activities such as online deliberations or

researcharxiv-cs-cl
21 Apr 2026
Research

The impact of postediting on AI generative translation in Yemeni context: Translating literary prose by ChatGPT

DGX agent

arXiv:2604.16704v1 Announce Type: new Abstract: This study examines the role of artificial intelligence in translation, focusing on ChatGPT, specifically ChatGPT-4, and the extent to which human poste

researcharxiv-cs-cl
21 Apr 2026
Research

Topology Structure Optimization of Reservoirs Using GLMY Homology

DGX agent

arXiv:2509.11612v3 Announce Type: replace Abstract: Reservoir is an efficient network for time series processing. It is well known that network structure is one of the determinants of its performance.

researcharxiv-cs-lg
21 Apr 2026
Tutorials

Towards a Foundation-Model Paradigm for Aerodynamic Prediction in Three-dimensional Design

DGX agent

arXiv:2604.18062v1 Announce Type: new Abstract: Accurate machine-learning models for aerodynamic prediction are essential for accelerating shape optimization, yet remain challenging to develop for com

tutorialsarxiv-cs-lg
21 Apr 2026
Research

Towards Reliable Testing of Machine Unlearning

DGX agent

arXiv:2604.16536v1 Announce Type: new Abstract: Machine learning components are now central to AI-infused software systems, from recommendations and code assistants to clinical decision support. As re

researcharxiv-cs-lg
21 Apr 2026
Agents

Towards Self-Improving Error Diagnosis in Multi-Agent Systems

DGX agent

arXiv:2604.17658v1 Announce Type: cross Abstract: Large Language Model (LLM)-based Multi-Agent Systems (MAS) enable complex problem-solving but introduce significant debugging challenges, characterize

agentsarxiv-cs-cl
21 Apr 2026
Safety

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts

DGX agent

arXiv:2604.18473v1 Announce Type: new Abstract: Extending a fully post-trained language model with new domain capabilities is fundamentally limited by monolithic training paradigms: retraining from sc

safetyarxiv-cs-lg
21 Apr 2026
Research

Understanding Counting Mechanisms in Large Language and Vision-Language Models

DGX agent

arXiv:2511.17699v2 Announce Type: replace Abstract: Counting is one of the fundamental abilities of large language models (LLMs) and large vision-language models (LVLMs). This paper examines how these

researcharxiv-cs-cv
21 Apr 2026
Research

Wasserstein-p Central Limit Theorem Rates: From Local Dependence to Markov Chains

DGX agent

arXiv:2601.08184v3 Announce Type: replace-cross Abstract: Non-asymptotic central limit theorem (CLT) rates play a central role in modern machine learning and operations research. In this paper, we stu

researcharxiv-cs-lg
21 Apr 2026
Research

Why Training-Free Token Reduction Collapses: The Inherent Instability of Pairwise Scoring Signals

DGX agent

arXiv:2604.16745v1 Announce Type: cross Abstract: Training-free token reduction methods for Vision Transformers (ToMe, ToFu, PiToMe, and MCTF) employ different scoring mechanisms, yet they share a clo

researcharxiv-cs-cv
21 Apr 2026
Model Releases

A PennyLane-Centric Dataset to Enhance LLM-based Quantum Code Generation using RAG

DGX agent

arXiv:2503.02497v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) offer powerful capabilities in code generation, natural language understanding, and domain-specific reasoning. Th

model-releasesarxiv-cs-ai
20 Apr 2026
Tutorials

Analyzing Chain of Thought (CoT) Approaches in Control Flow Code Deobfuscation Tasks

DGX agent

arXiv:2604.15390v1 Announce Type: cross Abstract: Code deobfuscation is the task of recovering a readable version of a program while preserving its original behavior. In practice, this often requires

tutorialsarxiv-cs-ai
20 Apr 2026
Research

Anthropomorphism and Trust in Human-Large Language Model interactions

DGX agent

arXiv:2604.15316v1 Announce Type: cross Abstract: With large language models (LLMs) becoming increasingly prevalent in daily life, so too has the tendency to attribute to them human-like minds and emo

researcharxiv-cs-ai
20 Apr 2026
Applications

Applied Explainability for Large Language Models: A Comparative Study

DGX agent

arXiv:2604.15371v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult t

applicationsarxiv-cs-ai
20 Apr 2026
Research

Author-in-the-Loop Response Generation and Evaluation: Integrating Author Expertise and Intent in Responses to Peer Review

DGX agent

arXiv:2602.11173v2 Announce Type: replace Abstract: Author response (rebuttal) writing is a critical stage of scientific peer review that demands substantial author effort. In practice, authors posses

researcharxiv-cs-cl
20 Apr 2026
Model Releases

Beyond Distribution Sharpening: The Importance of Task Rewards

DGX agent

arXiv:2604.16259v1 Announce Type: cross Abstract: Frontier models have demonstrated exceptional capabilities following the integration of task-reward-based reinforcement learning (RL) into their train

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Beyond Passive Viewing: A Pilot Study of a Hybrid Learning Platform Augmenting Video Lectures with Conversational AI

DGX agent

arXiv:2604.15334v1 Announce Type: cross Abstract: The exponential growth of AI education has brought millions of learners to online platforms, yet this massive scale has simultaneously exposed critica

applicationsarxiv-cs-ai
20 Apr 2026
Research

Bureaucratic Silences: What the Canadian AI Register Reveals, Omits, and Obscures

DGX agent

arXiv:2604.15514v1 Announce Type: new Abstract: In November 2025, the Government of Canada operationalized its commitment to transparency by releasing its first Federal AI Register. In this paper, we

researcharxiv-cs-ai
20 Apr 2026
Research

Can LLMs Understand the Impact of Trauma? Costs and Benefits of LLMs Coding the Interviews of Firearm Violence Survivors

DGX agent

arXiv:2604.16132v1 Announce Type: cross Abstract: Firearm violence is a pressing public health issue, yet research into survivors' lived experiences remains underfunded and difficult to scale. Qualita

researcharxiv-cs-ai
20 Apr 2026
Model Releases

ChatENV: An Interactive Vision-Language Model for Sensor-Guided Environmental Monitoring and Scenario Simulation

DGX agent

arXiv:2508.10635v3 Announce Type: replace Abstract: Understanding environmental changes from remote sensing imagery is vital for climate resilience, urban planning, and ecosystem monitoring. Yet, curr

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Enhancing Hazy Wildlife Imagery: AnimalHaze3k and IncepDehazeGan

DGX agent

arXiv:2604.16284v1 Announce Type: new Abstract: Atmospheric haze significantly degrades wildlife imagery, impeding computer vision applications critical for conservation, such as animal detection, tra

researcharxiv-cs-cv
20 Apr 2026
Model Releases

EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems

DGX agent

arXiv:2510.13220v2 Announce Type: replace Abstract: A fundamental limitation of current AI agents is their inability to learn complex skills on the fly at test time, often behaving like 'clever but cl

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

From Intention to Text: AI-Supported Goal Setting in Academic Writing

DGX agent

arXiv:2604.15800v1 Announce Type: cross Abstract: This study presents WriteFlow, an AI voice-based writing assistant designed to support reflective academic writing through goal-oriented interaction.

safetyarxiv-cs-ai
20 Apr 2026
Agents

From Multi-Agent to Single-Agent: When Is Skill Distillation Beneficial?

DGX agent

arXiv:2604.01608v2 Announce Type: replace Abstract: Multi-agent systems (MAS) tackle complex tasks by distributing expertise, though this often comes at the cost of heavy coordination overhead, contex

agentsarxiv-cs-ai
20 Apr 2026
Safety

How people use Copilot for Health

DGX agent

arXiv:2604.15331v1 Announce Type: cross Abstract: We analyze over 500,000 de-identified health-related conversations with Microsoft Copilot from January 2026 to characterize what people ask conversati

safetyarxiv-cs-ai
20 Apr 2026
Safety

Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding

DGX agent

arXiv:2512.04847v2 Announce Type: replace-cross Abstract: Pre-trained audio models excel at detecting acoustic patterns in auscultation sounds but often fail to grasp their clinical significance, limi

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

LLMs Corrupt Your Documents When You Delegate

DGX agent

arXiv:2604.15597v1 Announce Type: new Abstract: Large Language Models (LLMs) are poised to disrupt knowledge work, with the emergence of delegated work as a new interaction paradigm (e.g., vibe coding

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Long-Term Memory for VLA-based Agents in Open-World Task Execution

DGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

safetyarxiv-cs-ro
20 Apr 2026
Safety

M3R: Localized Rainfall Nowcasting with Meteorology-Informed MultiModal Attention

DGX agent

arXiv:2604.15377v1 Announce Type: cross Abstract: Accurate and timely rainfall nowcasting is crucial for disaster mitigation and water resource management. Despite recent advances in deep learning, pr

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition

DGX agent

arXiv:2604.16009v1 Announce Type: new Abstract: Metacognition, the ability to monitor and regulate one's own reasoning, remains under-evaluated in AI benchmarking. We introduce MEDLEY-BENCH, a benchma

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

NEFFY 2.0: A Breathing Companion Robot: User-Centered Design and Findings from a Study with Ukrainian Refugees

DGX agent

arXiv:2604.15325v1 Announce Type: cross Abstract: This paper presents the design of NEFFY 2.0, a social robot designed as a haptic slow-paced breathing companion for stress reduction, and reports find

model-releasesarxiv-cs-ro
20 Apr 2026
Tutorials

OT on the Map: Quantifying Domain Shifts in Geographic Space

DGX agent

arXiv:2604.16220v1 Announce Type: new Abstract: In computer vision and machine learning for geographic data, out-of-domain generalization is a pervasive challenge, arising from uneven global data cove

tutorialsarxiv-cs-lg
20 Apr 2026
Hardware

PyLO: Towards Accessible Learned Optimizers in PyTorch

DGX agent

arXiv:2506.10315v3 Announce Type: replace Abstract: Learned optimizers have been an active research topic over the past decade, with increasing progress toward practical, general-purpose optimizers th

hardwarearxiv-cs-lg
20 Apr 2026
Applications

SCRIPT: Implementing an Intelligent Tutoring System for Programming in a German University Context

DGX agent

arXiv:2604.16117v1 Announce Type: cross Abstract: Practice and extensive exercises are essential in programming education. Intelligent tutoring systems (ITSs) are a viable option to provide individual

applicationsarxiv-cs-ai
20 Apr 2026
Applications

Technically Love: The Evolution of Human-AI Romance Discourse on Reddit

DGX agent

arXiv:2604.15333v1 Announce Type: cross Abstract: Human-AI romantic relationships are increasingly common, yet little is understood about how public discourse around them emerges and shifts over time.

applicationsarxiv-cs-ai
20 Apr 2026
Research

Towards Rigorous Explainability by Feature Attribution

DGX agent

arXiv:2604.15898v1 Announce Type: new Abstract: For around a decade, non-symbolic methods have been the option of choice when explaining complex machine learning (ML) models. Unfortunately, such metho

researcharxiv-cs-ai
20 Apr 2026
Research

Verification Modulo Tested Library Contracts

DGX agent

arXiv:2604.15533v1 Announce Type: cross Abstract: We consider the problem of verification modulo tested library contracts as a step towards automating the verification of client programs that use comp

researcharxiv-cs-lg
20 Apr 2026
Research

VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck

DGX agent

arXiv:2601.05547v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in multimodal tasks, but remain susceptible to hallucinations, where gener

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Why Fine-Tuning Encourages Hallucinations and How to Fix It

DGX agent

arXiv:2604.15574v1 Announce Type: cross Abstract: Large language models are prone to hallucinating factually incorrect statements. A key source of these errors is exposure to new factual information t

model-releasesarxiv-cs-ai
20 Apr 2026
← Previous
1…100101102103104…108
Next →