AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,425 results
Safety

ZoFia: Zero-Shot Fake News Detection with Entity-Guided Retrieval and Multi-LLM Interaction

DGX agent

arXiv:2511.01188v2 Announce Type: replace Abstract: The rapid spread of fake news threatens social stability and public trust, highlighting the urgent need for its effective detection. Although large

safetyarxiv-cs-cl
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

A Comparative Study on the Impact of Traditional Learning and Interactive Learning on Students' Academic Performance and Emotional Well-Being

DGX agent

arXiv:2604.15335v1 Announce Type: cross Abstract: The growing adoption of interactive learning tools in higher education offers new opportunities to enhance student performance and well-being. This st

tutorialsarxiv-cs-ai
20 Apr 2026
Tutorials

A methodology to rank importance of frequencies and channels in electromyography data with Decision Tree classifiers

DGX agent

arXiv:2604.15353v1 Announce Type: cross Abstract: This study presents a methodology for identifying the most informative frequencies and channels in electromyography (EMG) data to evaluate muscle reco

tutorialsarxiv-cs-lg
20 Apr 2026
Applications

AI-assisted Protocol Information Extraction For Improved Accuracy and Efficiency in Clinical Trial Workflows

DGX agent

arXiv:2602.00052v2 Announce Type: replace-cross Abstract: Increasing clinical trial protocol complexity, amendments, and challenges around knowledge management create significant burden for trial team

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

AISysRev -- LLM-based Tool for Title-abstract Screening

DGX agent

arXiv:2510.06708v3 Announce Type: replace-cross Abstract: Conducting systematic reviews is laborious. In the screening or study selection phase, the number of papers can be overwhelming. Recent resear

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Applied Explainability for Large Language Models: A Comparative Study

DGX agent

arXiv:2604.15371v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult t

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

Beyond MCQ: An Open-Ended Arabic Cultural QA Benchmark with Dialect Variants

DGX agent

arXiv:2510.24328v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used to answer everyday questions, yet their performance on culturally grounded and dialectal co

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

Bilevel Optimization of Agent Skills via Monte Carlo Tree Search

DGX agent

arXiv:2604.15709v1 Announce Type: new Abstract: Agent exttt{skills} are structured collections of instructions, tools, and supporting resources that help large language model (LLM) agents perform part

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Co…

DGX agent

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Codex land near the human median, but with far tighter dispers

model-releasesethan-mollick--x
20 Apr 2026
Safety

Cognitive Agency Surrender: Defending Epistemic Sovereignty via Scaffolded AI Friction

DGX agent

arXiv:2603.21735v2 Announce Type: replace-cross Abstract: The proliferation of Generative Artificial Intelligence has transformed benign cognitive offloading into a systemic risk of cognitive agency s

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

CTSCAN: Evaluation Leakage in Chest CT Segmentation and a Reproducible Patient-Disjoint Benchmark

DGX agent

arXiv:2604.15561v1 Announce Type: cross Abstract: Reported chest CT segmentation performance can be strongly inflated when train and test partitions mix slices from the same study. We present CTSCAN,

model-releasesarxiv-cs-cv
20 Apr 2026
Applications

Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning

DGX agent

arXiv:2604.16029v1 Announce Type: new Abstract: Parallel reasoning enhances Large Reasoning Models (LRMs) but incurs prohibitive costs due to futile paths caused by early errors. To mitigate this, pat

applicationsarxiv-cs-cl
20 Apr 2026
Model Releases

DASB -- Discrete Audio and Speech Benchmark

DGX agent

arXiv:2406.14294v3 Announce Type: replace-cross Abstract: Discrete audio tokens have recently gained considerable attention for their potential to bridge audio and language processing, enabling multim

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

DataCenterGym: A Physics-Grounded Simulator for Multi-Objective Data Center Scheduling

DGX agent

arXiv:2604.15594v1 Announce Type: cross Abstract: Modern datacenters schedule heterogeneous workloads across geo-distributed sites with diverse compute capacities, electricity prices, and thermal cond

local-aiarxiv-cs-ai
20 Apr 2026
Local Ai

DINOv3 Beats Specialized Detectors: A Simple Foundation Model Baseline for Image Forensics

DGX agent

arXiv:2604.16083v1 Announce Type: new Abstract: With the rapid advancement of deep generative models, realistic fake images have become increasingly accessible, yet existing localization methods rely

local-aiarxiv-cs-cv
20 Apr 2026
Agents

Discover and Prove: An Open-source Agentic Framework for Hard Mode Automated Theorem Proving in Lean 4

DGX agent

arXiv:2604.15839v1 Announce Type: new Abstract: Most ATP benchmarks embed the final answer within the formal statement -- a convention we call 'Easy Mode' -- a design that simplifies the task relative

agentsarxiv-cs-ai
20 Apr 2026
Safety

DyTact: Capturing Dynamic Contacts in Hand-Object Manipulation

DGX agent

arXiv:2506.03103v2 Announce Type: replace Abstract: Reconstructing dynamic hand-object contacts is essential for realistic manipulation in AI character animation, XR, and robotics, yet it remains chal

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA …

DGX agent

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA that actually trains end-to-end from raw pixels. 15 million

model-releasesyann-lecun--x
20 Apr 2026
Model Releases

'Excuse me, may I say something...' CoLabScience, A Proactive AI Assistant for Biomedical Discovery and LLM-Expert Collaborations

DGX agent

arXiv:2604.15588v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into scientific workflows presents exciting opportunities to accelerate biomedical discovery. However,

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation

DGX agent

arXiv:2507.18756v2 Announce Type: replace Abstract: Multi-Armed Bandit (MAB) algorithms are widely used in recommender systems that require continuous, incremental learning. A core aspect of MABs is t

safetyarxiv-cs-lg
20 Apr 2026
Model Releases

Exploring the Capability Boundaries of LLMs in Mastering of Chinese Chouxiang Language

DGX agent

arXiv:2604.15841v1 Announce Type: new Abstract: While large language models (LLMs) have achieved remarkable success in general language tasks, their performance on Chouxiang Language, a representative

model-releasesarxiv-cs-cl
20 Apr 2026
Applications

From Articles to Canopies: Knowledge-Driven Pseudo-Labelling for Tree Species Classification using LLM Experts

DGX agent

arXiv:2604.16115v1 Announce Type: new Abstract: Hyperspectral tree species classification is challenging due to limited and imbalanced class labels, spectral mixing (overlapping light signatures from

applicationsarxiv-cs-cv
20 Apr 2026
Safety

From Intention to Text: AI-Supported Goal Setting in Academic Writing

DGX agent

arXiv:2604.15800v1 Announce Type: cross Abstract: This study presents WriteFlow, an AI voice-based writing assistant designed to support reflective academic writing through goal-oriented interaction.

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

From Zero to Detail: A Progressive Spectral Decoupling Paradigm for UHD Image Restoration with New Benchmark

DGX agent

arXiv:2604.15654v1 Announce Type: new Abstract: Ultra-high-definition (UHD) image restoration poses unique challenges due to the high spatial resolution, diverse content, and fine-grained structures p

model-releasesarxiv-cs-cv
20 Apr 2026
Tools

Important to note: that 3x increase for images is entirely due to Opus 4.7 being able to handle higher resolutions. I tried that again with …

DGX agent

Important to note: that 3x increase for images is entirely due to Opus 4.7 being able to handle higher resolutions. I tried that again with a 682x318 pixel image and it took 314 tokens with Opus 4.7 a

toolssimon-willison--x
20 Apr 2026
Agents

Integrating Graphs, Large Language Models, and Agents: Reasoning and Retrieval

DGX agent

arXiv:2604.15951v1 Announce Type: new Abstract: Generative AI, particularly Large Language Models, increasingly integrates graph-based representations to enhance reasoning, retrieval, and structured d

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

Interpretable Traces, Unexpected Outcomes: Investigating the Disconnect in Trace-Based Knowledge Distillation

DGX agent

arXiv:2505.13792v2 Announce Type: replace-cross Abstract: Recent advances in reasoning-focused Large Language Models (LLMs) have introduced Chain-of-Thought (CoT) traces - intermediate reasoning steps

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

IPQA: A Benchmark for Core Intent Identification in Personalized Question Answering

DGX agent

arXiv:2510.23536v2 Announce Type: replace Abstract: Intent identification serves as the foundation for generating appropriate responses in personalized question answering (PQA). However, existing benc

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des compo…

DGX agent

Je suis passé à Découverte de @CBCRadioCanada pour discuter des risques de l’IA, des raisons scientifiques qui expliquent certains des comportements inquiétants des modèles de pointe, et des solutions

safetyyoshua-bengio--x
20 Apr 2026
Tools

Kimi has been the most popular model on Fireworks, both out of the box and as a fine-tuning base (including Composer 2) Now Kimi K2.6 is liv…

DGX agent

Kimi has been the most popular model on Fireworks, both out of the box and as a fine-tuning base (including Composer 2) Now Kimi K2.6 is live with huge jumps (10+%) in coding, long-running agents and

toolsfireworks-ai--x
20 Apr 2026
Industry

Kimi-K2.6 is on HuggingFace

DGX agent

Kimi-K2.6 is a language model that has been released on HuggingFace, a popular platform for sharing machine learning models and datasets. The announcement was made by Clem Delangue, likely indicating

industryclem-delangue--x
20 Apr 2026
Safety

Language, Place, and Social Media: Geographic Dialect Alignment in New Zealand

DGX agent

arXiv:2604.15744v1 Announce Type: new Abstract: This thesis investigates geographic dialect alignment in place-informed social media communities, focussing on New Zealand-related Reddit communities. B

safetyarxiv-cs-cl
20 Apr 2026
Model Releases

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments

DGX agent

arXiv:2604.15384v1 Announce Type: cross Abstract: We introduce LinuxArena, a control setting in which agents operate directly on live, multi-service production environments. LinuxArena contains 20 env

model-releasesarxiv-cs-ai
20 Apr 2026
Tools

llm-openrouter 0.6

DGX agent

Release: llm-openrouter 0.6 llm openrouter refresh command for refreshing the list of available models without waiting for the cache to expire. I added this feature so I could try Kimi 2.6 on OpenRout

toolssimon-willison
20 Apr 2026
Model Releases

LLMSniffer: Detecting LLM-Generated Code via GraphCodeBERT and Supervised Contrastive Learning

DGX agent

arXiv:2604.16058v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Models (LLMs) in software development has made distinguishing AI-generated code from human-written code a cr

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

MM-Telco: Benchmarks and Multimodal Large Language Models for Telecom Applications

DGX agent

arXiv:2511.13131v2 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for automating complex reasoning and decision-making tasks. In telecommunications, they

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

Nice paper combining the strength of Skills and RAG. Most RAG systems retrieve on every query, whether the model needs help or not. This is …

DGX agent

Nice paper combining the strength of Skills and RAG. Most RAG systems retrieve on every query, whether the model needs help or not. This is wasteful when the model already knows the answer, and often

agentsdair-ai--x
20 Apr 2026
Model Releases

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation

DGX agent

arXiv:2604.15670v1 Announce Type: new Abstract: Reasoning segmentation has recently expanded from ground-level scenes to remote-sensing imagery, yet UAV data poses distinct challenges, including obliq

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Puppets or partners? Governing cyborg propaganda in the digital public square

DGX agent

arXiv:2602.13088v2 Announce Type: replace-cross Abstract: The distinction between genuine grassroots activism and automated influence operations is collapsing. While contemporary policy debates priori

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

RedBench: A Universal Dataset for Comprehensive Red Teaming of Large Language Models

DGX agent

arXiv:2601.03699v2 Announce Type: replace Abstract: As large language models (LLMs) become integral to safety-critical applications, ensuring their robustness against adversarial prompts is paramount.

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity

DGX agent

arXiv:2509.25897v2 Announce Type: replace-cross Abstract: People often encounter role conflicts -- social dilemmas where the expectations of multiple roles clash and cannot be simultaneously fulfilled

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Seed1.8 Model Card: Towards Generalized Real-World Agency

DGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval

DGX agent

arXiv:2604.15735v1 Announce Type: cross Abstract: Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-dr

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

TabularMath: Understanding Math Reasoning over Tables with Large Language Models

DGX agent

arXiv:2505.19563v4 Announce Type: replace Abstract: Mathematical reasoning has long been a key benchmark for evaluating large language models. Although substantial progress has been made on math word

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Technically Love: The Evolution of Human-AI Romance Discourse on Reddit

DGX agent

arXiv:2604.15333v1 Announce Type: cross Abstract: Human-AI romantic relationships are increasingly common, yet little is understood about how public discourse around them emerges and shifts over time.

applicationsarxiv-cs-ai
20 Apr 2026
Safety

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to…

DGX agent

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to the problem of how one could even control something vastly

safetyconnor-leahy--x
20 Apr 2026
Model Releases

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' compu…

DGX agent

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' computer use since @romainhuet demoed the ChatGPT App with 4o Vis

model-releasesswyx--x
20 Apr 2026
Applications

The Crutch or the Ceiling? How Different Generations of LLMs Shape EFL Student Writings

DGX agent

arXiv:2604.15460v1 Announce Type: cross Abstract: The rapid evolution of Large Language Models (LLMs) has made them powerful tools for enhancing student writing. This study explores the extent and lim

applicationsarxiv-cs-ai
20 Apr 2026
← Previous
1…519520521522523…530
Next →