AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,428 results
Model Releases

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

DGX agent

arXiv:2506.07180v3 Announce Type: replace-cross Abstract: As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reason

model-releasesarxiv-cs-ai
1 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Foreclassing: A new machine learning perspective on human decision making with temporal data

DGX agent

arXiv:2503.04956v2 Announce Type: replace-cross Abstract: Time series forecasts are widely used to inform decisions. Human decision-makers interpret these forecasts, incorporate prior experience and u

applicationsarxiv-cs-lg
1 May 2026
Model Releases

Function-based Parametric Co-Design Optimization of Dexterous Hands

DGX agent

arXiv:2604.27557v1 Announce Type: new Abstract: Despite advances in dexterous hand manipulation, robotic hand design is still largely decoupled from task-driven evaluation and control, limiting system

model-releasesarxiv-cs-ro
1 May 2026
Applications

GourNet: A CNN-Based Model for Mango Leaf Disease Detection

DGX agent

arXiv:2604.27764v1 Announce Type: new Abstract: Mango cultivation is crucial in the agricultural sector, significantly contributing to economic development and food security. However, diseases affecti

applicationsarxiv-cs-cv
1 May 2026
Tutorials

Graph World Models: Concepts, Taxonomy, and Future Directions

DGX agent

arXiv:2604.27895v1 Announce Type: new Abstract: As one of the mainstream models of artificial intelligence, world models allow agents to learn the representation of the environment for efficient predi

tutorialsarxiv-cs-ai
1 May 2026
Agents

GUI Agents with Reinforcement Learning: Toward Digital Inhabitants

DGX agent

arXiv:2604.27955v1 Announce Type: new Abstract: Graphical User Interface (GUI) agents have emerged as a promising paradigm for intelligent systems that perceive and interact with graphical interfaces

agentsarxiv-cs-ai
1 May 2026
Model Releases

GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance

DGX agent

arXiv:2503.12844v2 Announce Type: replace Abstract: For people affected by blindness and low vision (BLV), safe and independent navigation remains a major challenge, impacting over 2.2 billion individ

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats

DGX agent

arXiv:2604.27470v1 Announce Type: new Abstract: Millions of clinicians use ChatGPT to support clinical care, but evaluations of the most common use cases in model-clinician conversations are limited.

model-releasesarxiv-cs-cl
1 May 2026
Agents

here's a deep dive on how middleware lets you customize your agent harness, excellent writeup by @Vtrivedy10 !! deepagents offers a powerful…

DGX agent

here's a deep dive on how middleware lets you customize your agent harness, excellent writeup by @Vtrivedy10 !! deepagents offers a powerful base harness that you can customize for your use case! crea

agentsharrison-chase--x
1 May 2026
Model Releases

How Generative AI Disrupts Search: An Empirical Study of Google Search, Gemini, and AI Overviews

DGX agent

arXiv:2604.27790v1 Announce Type: cross Abstract: Generative AI is being increasingly integrated into web search for the convenience it provides users. In this work, we aim to understand how generativ

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

DGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

model-releasesdair-ai--x
1 May 2026
Model Releases

Improving Graph Few-shot Learning with Hyperbolic Space and Denoising Diffusion

DGX agent

arXiv:2604.27462v1 Announce Type: cross Abstract: Graph few-shot learning, which focuses on effectively learning from only a small number of labeled nodes to quickly adapt to new tasks, has garnered s

model-releasesarxiv-cs-ai
1 May 2026
Agents

It is honestly shocking how little code can get you so far within this base primitive. create_agent is one of the most fun things for me to …

DGX agent

It is honestly shocking how little code can get you so far within this base primitive. create_agent is one of the most fun things for me to show people who are trying to get started here - because it

agentsharrison-chase--x
1 May 2026
Model Releases

K2MUSE: A human lower-limb multimodal walking dataset spanning task and acquisition variability for rehabilitation robotics

DGX agent

arXiv:2504.14602v2 Announce Type: replace-cross Abstract: The natural interaction and control performance of lower limb rehabilitation robots are closely linked to biomechanical information from vario

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets

DGX agent

arXiv:2509.15549v2 Announce Type: replace Abstract: Multilingual instruction fine-tuning (IFT) empowers large language models to generalize across diverse linguistic and cultural contexts; however, hi

model-releasesarxiv-cs-cl
1 May 2026
Safety

METASYMBO: Multi-Agent Language-Guided Metamaterial Discovery via Symbolic Latent Evolution

DGX agent

arXiv:2604.27300v1 Announce Type: new Abstract: Metamaterial discovery seeks microstructured materials whose geometry induces targeted mechanical behavior. Existing inverse-design methods can efficien

safetyarxiv-cs-ai
1 May 2026
Local Ai

Multi-Level Narrative Evaluation Outperforms Lexical Features for Mental Health

DGX agent

arXiv:2604.27846v1 Announce Type: new Abstract: How people narrate their experiences offers a window into how the mind organizes them. Computational approaches to therapeutic writing have evolved from

local-aiarxiv-cs-cl
1 May 2026
Agents

One thing I love about LangChain is how the OSS pieces build on each other You can build robust workflows directly with LangGraph, our orche…

DGX agent

One thing I love about LangChain is how the OSS pieces build on each other You can build robust workflows directly with LangGraph, our orchestration framework. We also use it as the foundation for Dee

agentsharrison-chase--x
1 May 2026
Model Releases

OR-VSKC: Resolving Visual-Semantic Knowledge Conflicts in Operating Rooms with Synthetic Data-Guided Alignment

DGX agent

arXiv:2506.22500v2 Announce Type: replace-cross Abstract: Automated identification of surgical safety risks is critical for improving patient outcomes; however, Multimodal Large Language Models (MLLMs

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Parameter-Efficient Architectural Modifications for Translation-Invariant CNNs

DGX agent

arXiv:2604.27870v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) are widely assumed to be translation-invariant, yet standard architectures exhibit a startling fragility: even a si

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Physical Foundation Models: Fixed hardware implementations of large-scale neural networks

DGX agent

arXiv:2604.27911v1 Announce Type: new Abstract: Foundation models are deep neural networks (such as GPT-5, Gemini~3, and Opus~4) trained on large datasets that can perform diverse downstream tasks --

model-releasesarxiv-cs-lg
1 May 2026
Safety

Political Bias Audits of LLMs Capture Sycophancy to the Inferred Auditor

DGX agent

arXiv:2604.27633v1 Announce Type: new Abstract: Large language models (LLMs) are commonly evaluated for political bias based on their responses to fixed questionnaires, which typically place frontier

safetyarxiv-cs-ai
1 May 2026
Agents

Pragmos: A Process Agentic Modeling System

DGX agent

arXiv:2604.27311v1 Announce Type: cross Abstract: The advent of Large Language Models (LLMs) has significantly transformed tasks across Software Engineering. In the context of Business Process Managem

agentsarxiv-cs-ai
1 May 2026
Model Releases

REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)

DGX agent

arXiv:2604.27319v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved remarkable progress in recent years, driving their adoption across a wide range of domains, including compu

model-releasesarxiv-cs-lg
1 May 2026
Agents

Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study

DGX agent

arXiv:2604.27464v1 Announce Type: cross Abstract: Autonomous agent frameworks built upon large language models (LLMs) are evolving into complex, tool-integrated, and continuously operating systems, in

agentsarxiv-cs-ai
1 May 2026
Agents

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation

DGX agent

arXiv:2604.27620v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) aims to enable an embodied agent to follow natural-language instructions and navigate to a target location in unsee

agentsarxiv-cs-cv
1 May 2026
Model Releases

SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images

DGX agent

arXiv:2604.28039v1 Announce Type: new Abstract: Spectra are a prevalent yet highly information-dense form of scientific imagery, presenting substantial challenges to multimodal large language models (

model-releasesarxiv-cs-ai
1 May 2026
Industry

Standard Intelligence raises $75M to develop efficient computer use models

DGX agent

Standard Intelligence Inc., a six-person artificial intelligence startup, today announced that it has raised 75 million in funding. Sequoia and Spark Capital led the round. They were joined by multipl

industrysiliconangle
1 May 2026
Model Releases

The TEA Nets framework combines AI and cognitive network science to model targets, events and actors in text

DGX agent

arXiv:2604.27673v1 Announce Type: new Abstract: We introduce Target-Event-Agent Networks (TEA Nets) as a computational framework to extract subjects (``Agents'), verbs (``Events'), and objects (``Targ

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Towards All-Day Perception for Off-Road Driving: A Large-Scale Multispectral Dataset and Comprehensive Benchmark

DGX agent

arXiv:2604.27499v1 Announce Type: new Abstract: Off-road nighttime autonomous driving suffers from unreliable visible-light perception, making infrared modality crucial for accurate freespace detectio

model-releasesarxiv-cs-cv
1 May 2026
Safety

Towards Neuro-symbolic Causal Rule Synthesis, Verification, and Evaluation Grounded in Legal and Safety Principles

DGX agent

arXiv:2604.28087v1 Announce Type: cross Abstract: Rule-based systems remain central in safety-critical domains but often struggle with scalability, brittleness, and goal misspecification. These limita

safetyarxiv-cs-ai
1 May 2026
Applications

TwinGate: Stateful Defense against Decompositional Jailbreaks in Untraceable Traffic via Asymmetric Contrastive Learning

DGX agent

arXiv:2604.27861v1 Announce Type: cross Abstract: Decompositional jailbreaks pose a critical threat to large language models (LLMs) by allowing adversaries to fragment a malicious objective into a seq

applicationsarxiv-cs-cl
1 May 2026
Applications

VERA: Generating Visual Explanations of Two-Dimensional Embeddings via Region Annotation

DGX agent

arXiv:2406.04808v2 Announce Type: replace Abstract: Two-dimensional embeddings obtained from dimensionality reduction techniques such as MDS, t-SNE, or UMAP, are widely used to visualize high-dimensio

applicationsarxiv-cs-lg
1 May 2026
Model Releases

Visual Analysis of Multi-outcome Causal Graphs

DGX agent

arXiv:2408.02679v3 Announce Type: replace Abstract: We introduce a visual analysis method for multiple causal graphs with different outcome variables, namely, multi-outcome causal graphs. Multi-outcom

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

What Makes a Good Terminal-Agent Benchmark Task: A Guideline for Adversarial, Difficult, and Legible Evaluation Design

DGX agent

arXiv:2604.28093v1 Announce Type: new Abstract: Terminal-agent benchmarks have become a primary signal for measuring the coding and system-administration capabilities of large language models. As the

model-releasesarxiv-cs-ai
1 May 2026
Applications

World2Minecraft: Occupancy-Driven Simulated Scenes Construction

DGX agent

arXiv:2604.27578v1 Announce Type: new Abstract: Embodied intelligence requires high-fidelity simulation environments to support perception and decision-making, yet existing platforms often suffer from

applicationsarxiv-cs-cv
1 May 2026
Safety

A Scaled Three-Vehicle Platooning Platform

DGX agent

arXiv:2604.25963v1 Announce Type: new Abstract: Vehicle platooning has attracted increasing attention as a promising approach to improve traffic efficiency, energy consumption, and roadway safety thro

safetyarxiv-cs-ro
30 Apr 2026
Agents

A Survey of Multi-Agent Deep Reinforcement Learning with Graph Neural Network-Based Communication

DGX agent

arXiv:2604.25972v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), the integration of a communication mechanism, allowing agents to better learn to coordinate their action

agentsarxiv-cs-ai
30 Apr 2026
Safety

A Survey of Process Reward Models: From Outcome Signals to Process Supervisions for Large Language Models

DGX agent

arXiv:2510.08049v3 Announce Type: replace-cross Abstract: Although Large Language Models (LLMs) exhibit advanced reasoning ability, conventional alignment remains largely dominated by outcome reward m

safetyarxiv-cs-ai
30 Apr 2026
Model Releases

Anthropic announces Claude Security public beta to find and fix software vulnerabilities

DGX agent

Anthropic PBC announced the launch of Claude Security in public beta mode today to help cybersecurity teams scan their codebases for vulnerabilities and generate patches. Part of Claude Enterprise, th

model-releasessiliconangle
30 Apr 2026
Model Releases

Anthropic unveils BioMysteryBench to test Claude's bioinformatics skills against human experts, and says Mythos solved ~30% of 23 questions that stumped experts (Anthropic)

DGX agent

Anthropic: Anthropic unveils BioMysteryBench to test Claude's bioinformatics skills against human experts, and says Mythos solved ~30% of 23 questions that stumped experts — In this post, Brianna, a r

model-releasestechmeme
30 Apr 2026
Model Releases

Breaking the Rigid Prior: Towards Articulated 3D Anomaly Detection

DGX agent

arXiv:2604.26868v1 Announce Type: new Abstract: Existing 3D anomaly detection methods are built on a rigid prior: normal geometry is pose-invariant and can be canonicalized through registration or ali

model-releasesarxiv-cs-cv
30 Apr 2026
Safety

Classification of Public Opinion on the Free Nutritional Meal Program on YouTube Media Using the LSTM Method

DGX agent

arXiv:2604.26312v1 Announce Type: new Abstract: Public opinion towards the Free Nutritious Meal Program (MBG) on YouTube social media reflects diverse community responses. This study applies the Long

safetyarxiv-cs-cl
30 Apr 2026
Model Releases

DB-KSVD: Scalable Alternating Optimization for Disentangling High-Dimensional Embedding Spaces

DGX agent

arXiv:2505.18441v2 Announce Type: replace Abstract: Dictionary learning has recently emerged as a promising approach for mechanistic interpretability of large transformer models. Disentangling high-di

model-releasesarxiv-cs-lg
30 Apr 2026
Safety

Evaluating the Alignment Between GeoAI Explanations and Domain Knowledge in Satellite-Based Flood Mapping

DGX agent

arXiv:2604.26051v1 Announce Type: cross Abstract: The increasing number of satellites has improved the temporal resolution of Earth observation, making satellite-based flood mapping a promising approa

safetyarxiv-cs-ai
30 Apr 2026
Safety

Generative Bid Shading in Real-Time Bidding Advertising

DGX agent

arXiv:2508.06550v3 Announce Type: replace-cross Abstract: Bid shading plays a crucial role in Real-Time Bidding (RTB) by adaptively adjusting the bid to avoid advertisers overspending. Existing mainst

safetyarxiv-cs-lg
30 Apr 2026
Model Releases

GIFGuard: Proactive Forensics against Deepfakes in Facial GIFs via Spatiotemporal Watermarking

DGX agent

arXiv:2604.26519v1 Announce Type: new Abstract: The rapid evolution of deepfake technology poses an unprecedented threat to the authenticity of Graphics Interchange Format (GIF) imagery, which serves

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

HER: Human-like Reasoning and Reinforcement Learning for LLM Role-playing

DGX agent

arXiv:2601.21459v4 Announce Type: replace-cross Abstract: LLM role-playing, i.e., using LLMs to simulate specific personas, has emerged as a key capability in various applications, such as companionsh

model-releasesarxiv-cs-ai
30 Apr 2026
← Previous
1…510511512513514…530
Next →