AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Safety

TROPT: An Open Framework for Unifying and Advancing Discrete Text Optimization

DGX agent

arXiv:2606.23496v1 Announce Type: new Abstract: Discrete text-trigger optimization -- searching for text sequences that, when ingested by a model, steer it toward a specified objective -- underpins mo

safetyarxiv-cs-lg
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

What Accuracy and Gradient Cosine Miss: Evaluating Feedback Alignment via Scale Stability, Reference Validity, and Depth Utility

DGX agent

arXiv:2606.21126v1 Announce Type: new Abstract: Despite the success of deep learning, training deep networks in biologically plausible and hardware-efficient ways remains an open challenge. Feedback a

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

A Comprehensive Ecosystem for Open-Domain Customized Video Generation

DGX agent

arXiv:2606.11783v1 Announce Type: new Abstract: Recent progress in video generation has shown impressive visual synthesis capabilities. However, open-domain customized video generation remains limited

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

A PubMed-Scale Dataset of Structured Biomedical Abstracts

DGX agent

arXiv:2606.11361v1 Announce Type: cross Abstract: Structured abstracts are important for biomedical literature processing, by facilitating information retrieval, text mining, and knowledge synthesis.

model-releasesarxiv-cs-cl
11 Jun 2026
Applications

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data

DGX agent

arXiv:2601.17717v3 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for generating data across various modalities. By transforming data from a scarce resour

applicationsarxiv-cs-ai
11 Jun 2026
Local Ai

A Turbo-Inference Strategy for Object Detection and Instance Segmentation

DGX agent

arXiv:2606.12371v1 Announce Type: new Abstract: Object detection and instance segmentation tasks are closely related. Existing top-down instance segmentation methods usually follow a detect-then-segme

local-aiarxiv-cs-cv
11 Jun 2026
Model Releases

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

DGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

DGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

AI Coding Agents Can Reproduce Social Science Findings

DGX agent

arXiv:2606.11447v1 Announce Type: new Abstract: Recent anecdotal evidence suggests that AI coding agents can reproduce published findings when provided with original data and code; yet systematic eval

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

An Ontology-Guided Multi-Anchor Graph Retrieval Framework for Traffic Legal Liability Determination

DGX agent

arXiv:2606.11910v1 Announce Type: new Abstract: Traffic law liability determination is critical for assigning legal penalties, requiring the simultaneous identification of interdependent statutory pro

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

AnchorEdit: Maintaining Temporal Consistency in Multi-turn Image Editing via Causal Memory

DGX agent

arXiv:2606.11751v1 Announce Type: cross Abstract: Multi-turn image editing is essential for iterative design, yet current models often struggle with identity drift and error accumulation over successi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Atlas H&E-TME: Scalable AI-Based Tissue Profiling at Expert Pathologist-Level Accuracy

DGX agent

arXiv:2606.12346v1 Announce Type: cross Abstract: Hematoxylin and eosin (H&E) staining is the cornerstone of histopathology, yet scalable, quantitative analysis of H&E whole-slide images (WSIs) remain

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

Automated Creativity Evaluation of Language Models Across Open-Ended Tasks

DGX agent

arXiv:2606.11762v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in language understanding, reasoning, and generation, sparking growing interest in thei

agentsarxiv-cs-ai
11 Jun 2026
Safety

Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework

DGX agent

arXiv:2606.12065v1 Announce Type: new Abstract: Automating compliance check for geometry-intensive regulations remains a significant technical bottleneck in Building Information Modeling (BIM), primar

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Benchmarking Cross-Domain Audio-Visual Deception Detection

DGX agent

arXiv:2405.06995v4 Announce Type: replace-cross Abstract: Automated deception detection is crucial for assisting humans in accurately assessing truthfulness and identifying deceptive behavior. Convent

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Benchmarking Large Language Models for Safety Data Extraction

DGX agent

arXiv:2606.11204v1 Announce Type: new Abstract: Accurate extraction of structured information from Safety Data Sheets (SDS) remains challenging in industrial safety due to heterogeneous document forma

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Can AI Agents Synthesize Scientific Conclusions?

DGX agent

arXiv:2606.11337v1 Announce Type: new Abstract: Scientific AI agents increasingly retrieve evidence, reason across sources, and synthesize conclusions used in consequential decisions. Yet, their abili

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

CCKS: Consensus-based Communication and Knowledge Sharing

DGX agent

arXiv:2606.12281v1 Announce Type: cross Abstract: In Decentralized Training and Decentralized Execution (DTDE) for cooperative Multi-Agent Reinforcement Learning (MARL), action-advising-based knowledg

agentsarxiv-cs-ai
11 Jun 2026
Hardware

Characterizing Software Aging in GPU-Based LLM Serving Systems

DGX agent

arXiv:2606.11916v1 Announce Type: cross Abstract: This paper proposes an empirical methodology to study software aging in GPU-based LLM serving systems. Traditional aging studies focus on CPU-centric

hardwarearxiv-cs-ai
11 Jun 2026
Model Releases

Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models

DGX agent

arXiv:2606.11324v1 Announce Type: cross Abstract: We introduce Embodied-R1.5, a unified Embodied Foundation Model (EFM) that integrates comprehensive embodied reasoning capabilities, spanning embodied

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

Fast-SDE: Efficient Single-Microphone Sound Source Distance Estimation in Reverberant Environments

DGX agent

arXiv:2606.12339v1 Announce Type: cross Abstract: Sound source distance estimation (SDE) is a critical capability in human-robot interaction. An inappropriate interaction distance not only reduces the

applicationsarxiv-cs-ro
11 Jun 2026
Model Releases

FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback

DGX agent

arXiv:2601.04203v2 Announce Type: replace Abstract: We present FronTalk, a benchmark for front-end code generation that pioneers the study of a unique interaction dynamic: conversational code generati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

I Understand How You Feel: Enhancing Deeper Emotional Support Through Multilingual Emotional Validation in Dialogue System

DGX agent

arXiv:2606.11875v1 Announce Type: new Abstract: Emotional validation - explicitly acknowledging that a user's feelings make sense - has proven therapeutic value but has received little computational a

model-releasesarxiv-cs-cl
11 Jun 2026
Applications

LaQual: An Automated Framework for LLM App Quality Evaluation

DGX agent

arXiv:2508.18636v2 Announce Type: replace-cross Abstract: Representing a new paradigm in software distribution, LLM app stores are rapidly emerging, offering users diverse choices for content generati

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

LifeSentence: Language models can encode human life course trajectories from longitudinal panel data

DGX agent

arXiv:2606.11220v1 Announce Type: new Abstract: Forecasting human life outcomes is important to gain insights into how individuals attain long and healthy lives. Conventional statistical approaches yi

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

LLMs+Graphs: Toward Graph-Native, Synergistic AI Systems

DGX agent

arXiv:2606.11560v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced rapidly, but their limitations in structured and multi-hop reasoning underscore the need for graph-native,

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

Modelling magnetic material properties with uncertainty-aware neural networks

DGX agent

arXiv:2606.11870v1 Announce Type: cross Abstract: Machine learning is increasingly applied to accelerate the discovery of novel materials by exploring large compositional and structural design spaces.

model-releasesarxiv-cs-lg
11 Jun 2026
Agents

Physics-informed generative AI for semiconductor manufacturing: Enforcing hard physical constraints in generative models by construction

DGX agent

arXiv:2606.11247v1 Announce Type: cross Abstract: Generative models are increasingly used to propose designs, data, and control actions for physical systems, yet many such systems are governed by hard

agentsarxiv-cs-ai
11 Jun 2026
Tutorials

Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!

DGX agent

arXiv:2504.09762v4 Announce Type: replace Abstract: Intermediate token generation (ITG), where a model produces output before the solution, has become a standard method to improve the performance of l

tutorialsarxiv-cs-ai
11 Jun 2026
Model Releases

Reassessing High-Performing LLMs on Polish Medical Exams: True Competence or Bias-Driven Performance?

DGX agent

arXiv:2606.12250v1 Announce Type: new Abstract: Large language models (LLMs) in medicine are mainly evaluated using multiple-choice question answering (MCQA), which can overestimate real clinical abil

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Reinforcement Learning Disrupts Gradient-Based Adversarial Optimization

DGX agent

arXiv:2606.12251v1 Announce Type: cross Abstract: Gradient-based adversarial attacks remain a dominant threat to deep neural networks (DNNs), as they exploit gradient information to efficiently optimi

safetyarxiv-cs-ai
11 Jun 2026
Safety

Semantically-Aware Diver Activity Recognition Framework for Effective Underwater Multi-Human-Robot Collaboration

DGX agent

arXiv:2606.12374v1 Announce Type: cross Abstract: Effective multi-human-robot collaboration is essential for expanding human-led operations in the challenging and high-risk underwater environment. For

safetyarxiv-cs-cv
11 Jun 2026
Model Releases

System Report for CCL25-Eval Task 5: New Dataset and LoRA-Fine-Tuned Qwen2.5

DGX agent

arXiv:2606.12392v1 Announce Type: cross Abstract: Recently, large language models (LLMs) have achieved promising progress in the fields of classical Chinese translation and the generation of classical

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

T2MM: An LLM Supported Architecture For Inquiry-Based Modeling

DGX agent

arXiv:2606.11210v1 Announce Type: cross Abstract: Model Construction is a foundational practice in science learning that relies on visualization and interactivity. Large Language Models, increasingly

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

Tac-DINO: Learning Vision-Tactile Features with Patch Alignment

DGX agent

arXiv:2606.12069v1 Announce Type: new Abstract: Touch is the primary medium through which humans interact with the environment. Currently, tactile learning mainly focuses on image-level pretraining or

model-releasesarxiv-cs-cv
11 Jun 2026
Applications

'That's AI Slop, You Bot!' Studying Accusations, Evidence, and Credibility in Online Discourse Towards LLM-Generated Comments

DGX agent

arXiv:2606.12073v1 Announce Type: cross Abstract: Generative AI has made fluent prose cheap to produce, breaking the old promise to readers that good writing meant real thinking. How have readers resp

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

The Language You Ask In: Language-Conditioned Ideological Divergence in LLM Analysis of Contested Political Documents

DGX agent

arXiv:2601.12164v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as analytical tools across multilingual contexts, yet their outputs may carry systemati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

DGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation

DGX agent

arXiv:2601.03792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable proficiency in general medical domains. However, their performance significantly degrades

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

A fine-grained attention and geometric correspondence model for musculoskeletal risk classification in athletes using multimodal visual and skeletal features

DGX agent

arXiv:2509.05913v3 Announce Type: replace Abstract: Musculoskeletal disorders pose significant risks to athletes, and early risk assessment is essential for prevention. However, most existing methods

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

A Large Scale Open-Source Image and Video Dataset for Robust Wildfire Detection and Classification

DGX agent

arXiv:2606.10174v1 Announce Type: new Abstract: Wildfire detection and monitoring are critical for mitigating fire spread and reducing environmental and infrastructural damage. In this work, we introd

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI

DGX agent

arXiv:2505.01458v2 Announce Type: replace-cross Abstract: Navigation and manipulation are core capabilities in Embodied AI, but training agents to perform them directly in the real world is costly, ti

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

A Unified Multi-Modal Framework for Intelligent Financial Systems: Integrating Reinforcement Learning, High-Frequency Trading, and Game-Theoretic Approaches with Cross-Modal Sentiment Analysis

DGX agent

arXiv:2606.10412v1 Announce Type: new Abstract: The rapid evolution of financial technology demands sophisticated artificial intelligence systems capable of handling diverse challenges across multiple

safetyarxiv-cs-ai
10 Jun 2026
Safety

Automated Alignment between Elicitation Interviews and Requirements

DGX agent

arXiv:2510.08622v2 Announce Type: replace Abstract: Software requirements are derived from a variety of elicitation techniques, many of which have a conversational nature, like interviews. However, ev

safetyarxiv-cs-cl
10 Jun 2026
Safety

Automated Scoring of Arabic Text Using Large Language Models: A Literature Review

DGX agent

arXiv:2606.09830v1 Announce Type: new Abstract: In modern educational systems, Automatic Text Scoring (ATS) plays a central role by enabling scalable and consistent evaluation of learner responses wit

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

Benchmarking and Exploring the Capabilities of LLMs for Attack Investigations

DGX agent

arXiv:2606.10281v1 Announce Type: cross Abstract: This paper presents AuditBench, a new benchmark dataset for evaluating the capabilities of LLMs at investigating security-related system audit logs. W

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts

DGX agent

arXiv:2606.10061v1 Announce Type: new Abstract: Large language models (LLMs) increasingly participate in emotionally sensitive social conversations, where responses may shift from balanced support tow

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Beyond Model Size: Probing the Gaps in Visual in-Context Learning by Training a Tiny Model

DGX agent

arXiv:2606.10905v1 Announce Type: new Abstract: Visual in-Context Learning (VICL) aims at making progress towards adaptive vision models, that can -- based on a few examples -- adapt to a new task at

model-releasesarxiv-cs-cv
10 Jun 2026
← Previous
1…424425426427428…462
Next →