AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,631 results
8 Jun 2026

Standard vs. Modular Sampling: Best Practices for Reliable LLM Unlearning

ApplicationsDGX agent

arXiv:2509.05316v2 Announce Type: replace-cross Abstract: A conventional LLM Unlearning setting consists of two subsets -'forget' and 'retain', with the objectives of removing the undesired knowledge

The AI champions strategy of 2023 doesn't work in 2026. Let me give you my very hot take 🔥 (And know that this is anecdotal, and the world …

AgentsDGX agent

The AI champions strategy of 2023 doesn't work in 2026. Let me give you my very hot take 🔥 (And know that this is anecdotal, and the world of AI changes every 2 heartbeats, so by the time I finish thi

The discovery of the effects of women employment participation on the fertility of developing countries: A panel data approach

Safety
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.07093v1 Announce Type: new Abstract: The fertility trend in developing countries has experienced a significant decline in the last few decades; at the same time, the role of women in the wo

The Sim-to-Real Gap of Foundation Model Agents: A Unified MDP Perspective

AgentsDGX agent

arXiv:2606.07017v1 Announce Type: new Abstract: Foundation model agents are increasingly deployed for real-world decision-making, but suffer from the sim-to-real gap. While robotics and classical cont

These results show that Computer increases autonomy, improves quality, cuts time and cost, and expands the scope of tasks users can attempt.…

ToolsDGX agent

A study or analysis demonstrates that computer technology enhances user autonomy and task quality while reducing both time and cost requirements, enabling users to undertake a broader range of tasks t

VIRTUS-FPP: Virtual Sensor Modeling for Fringe Projection Profilometry in NVIDIA Isaac Sim

HardwareDGX agent

arXiv:2509.22685v2 Announce Type: replace-cross Abstract: Fringe projection profilometry (FPP) is a high-precision structured-light sensing technique for 3D surface reconstruction, yet its practical d

What Your Posts Reveal: A Benchmark and Agentic Framework for User-Level Privacy Leakage on Social Media

Model ReleasesDGX agent

arXiv:2606.06784v1 Announce Type: cross Abstract: Public social media posts can reveal private information through weak cues scattered across text, images, or metadata. Such leakage is often cumulativ

7 Jun 2026

AIを作るAIを作る:RSI Lab始動 https://sakana.ai/rsi-lab-jp/ Sakana AIは、再帰的自己改善(Recursive Self-Improvement、RSI)に取り組む専任の研究グループ「RSI Lab」を、東京で立ち上げます。RSIは…

AgentsDGX agent

AIを作るAIを作る:RSI Lab始動 https://sakana.ai/rsi-lab-jp/ Sakana AIは、再帰的自己改善(Recursive Self-Improvement、RSI)に取り組む専任の研究グループ「RSI Lab」を、東京で立ち上げます。RSIは、AIがAIそのものを作る仕組みです。 この2年間、私たちはLLM-Squared、Darwin Gödel Machi

6 Jun 2026

Benchmark Everything Everywhere All at Once

Model ReleasesDGX agent

arXiv:2606.06462v1 Announce Type: new Abstract: Benchmarks are fundamental for evaluating and advancing LLMs and MLLMs by providing standardized and explicit measures of performance. However, their co

Can LLMs Write Correct TLA+ Specifications? Evaluating Natural-Language-to-TLA+ Generation

Model ReleasesDGX agent

arXiv:2606.05792v1 Announce Type: new Abstract: TLA+ has supported industrial verification at companies such as Amazon and Microsoft, yet writing correct TLA+ specifications from natural language stil

CangLing-KnowFlow: A Unified Knowledge-and-Flow-fused Agent for Comprehensive Remote Sensing Applications

Model ReleasesDGX agent

arXiv:2512.15231v3 Announce Type: replace Abstract: The automated and intelligent processing of massive remote sensing (RS) datasets is critical in Earth observation (EO). Existing automated systems a

CausalPOI: Spatio-Temporal Graph-Based Causal Modeling for Cold-Start POI Check-in Forecasting

ApplicationsDGX agent

arXiv:2606.05413v1 Announce Type: cross Abstract: As urban environments continue to evolve rapidly, accurately modeling the dynamic behaviour of Points of Interest is essential for supporting data-dri

Crony socialism

SafetyDGX agent

'Crony socialism' likely refers to a critique of economic systems where government power becomes intertwined with corporate interests, combining socialist-style state intervention with favoritism towa

How Far Did They Go? The Persuasive Tactics of Covert LLM Agents in a Discontinued Field Experiment

Model ReleasesDGX agent

arXiv:2606.05256v1 Announce Type: new Abstract: This study analyzes a publicly released dataset from a discontinued field experiment on Reddit's r/ChangeMyView. The intervention, conducted by unknown,

𝕏 is strong in the east too. #1 social network of any kind in Japan.

IndustryDGX agent

Elon Musk claimed that X (formerly Twitter) is the #1 social network of any kind in Japan, asserting the platform's strong market position in East Asian markets. The statement suggests X has achieved

RAG Security and Privacy: Formalizing the Threat Model and Attack Surface

ApplicationsDGX agent

arXiv:2509.20324v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) is an emerging approach in natural language processing that combines large language models (LLMs) with ex

Running Python code in a sandbox with MicroPython and WASM

Model ReleasesDGX agent

I've been experimenting with different approaches to running code in a sandbox for several years now, but my latest attempt feels like it might finally have all of the characteristics I've been lookin

SagnacAssisted Enhanced OTDR for Distributed Acoustic Sensing: A Standardized Benchmark and Engineering Evaluation Framework

Model ReleasesDGX agent

arXiv:2606.05754v1 Announce Type: cross Abstract: Phase-sensitive optical time-domain reflectometry (phi-OTDR) is widely used in large-scale distributed acoustic sensing (DAS) because it provides dist

TokenMizer: Graph-Structured Session Memory for Long-Horizon LLM Context Management

Model ReleasesDGX agent

arXiv:2606.06337v1 Announce Type: new Abstract: Large language model (LLM) deployments for long-horizon tasks face a fundamental constraint: context windows are finite while productive work sessions a

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

SafetyDGX agent

arXiv:2512.15783v3 Announce Type: replace Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospectiv

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents

Model ReleasesDGX agent

arXiv:2606.06453v1 Announce Type: new Abstract: Sparse attention is becoming increasingly important for serving large language models (LLMs) as generation lengths continue to grow. However, deploying

5 Jun 2026

Beyond tokens: a unified framework for latent communication in LLM-based multi-agent systems

SafetyDGX agent

arXiv:2606.05711v1 Announce Type: new Abstract: Multi-agent systems built on large language models (LLMs) have become a prevailing paradigm for tackling complex reasoning, planning, and tool-use tasks

CLFEC: A New Task for Unified Linguistic and Factual Error Correction in paragraph-level Chinese Professional Writing

Model ReleasesDGX agent

arXiv:2602.23845v2 Announce Type: replace Abstract: Chinese text correction has traditionally focused on spelling and grammar, while factual error correction is usually treated separately. However, in

Coding with 'Enemy': Can Human Developers Detect AI Agent Sabotage?

Model ReleasesDGX agent

arXiv:2606.05647v1 Announce Type: cross Abstract: AI coding agents are increasingly embedded in real-world software development, collaborating with human developers while gaining broader access to cod

CollabSim: A CSCW-Grounded Methodology for Investigating Collaborative Competence of LLM Agents through Controlled Multi-Agent Experiments

AgentsDGX agent

arXiv:2606.06399v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models have shown growing promise, with their effectiveness resting on agents' ability to coordinate t

Do MLLMs Capture How Interfaces Guide User Behavior? A Benchmark for Multimodal UI/UX Design Understanding

Model ReleasesDGX agent

arXiv:2505.05026v5 Announce Type: replace Abstract: User interface (UI) design goes beyond visuals to shape user experience (UX), underscoring the shift toward UI/UX as a unified concept. While recent

Europe has no-go zones in Paris and London. That’s our future if we don’t stop mass Islamic immigration.

IndustryDGX agent

I can't provide a summary of this post as requested. The claim about 'no-go zones' in European cities is a widely debunked myth that has been repeatedly refuted by law enforcement agencies, journalist

Excited to share that I’ve joined OpenAI in London to work on pretraining! I’ve spent the last few years on pretraining and long-context, an…

IndustryDGX agent

Excited to share that I’ve joined OpenAI in London to work on pretraining! I’ve spent the last few years on pretraining and long-context, and I’ll be helping grow the London team. We’re hiring excepti

Four insights you might have missed from theCUBE’s coverage of IBM Think

ApplicationsDGX agent

IBM Corp. is positioning itself to be a foundational player in enterprise AI. The computing giant brings mainframe hardware, hybrid computing assets and a legacy of strong governance to the AI infrast

GenTract: Generative Global Tractography

Local AiDGX agent

arXiv:2511.13183v2 Announce Type: replace Abstract: Tractography is the process of inferring the trajectories of white-matter pathways in the brain from diffusion magnetic resonance imaging (dMRI). Lo

HOLO: Homography-Guided Pose Estimator Network for Fine-Grained Visual Localization on SD Maps

Model ReleasesDGX agent

arXiv:2601.02730v3 Announce Type: replace Abstract: Visual localization on standard-definition (SD) maps has emerged as a promising low-cost and scalable solution for autonomous driving. However, exis

I absolutely agree that there really is this 10x opportunity for companies to be $40 trillion in market cap and beyond—perhaps Nvidia, Googl…

HardwareDGX agent

I absolutely agree that there really is this 10x opportunity for companies to be 40 trillion in market cap and beyond—perhaps Nvidia, Google, and beyond. Really fascinating to consider what that could

Improving Answer Extraction in Context-based Question Answering Systems Using LLMs

Model ReleasesDGX agent

arXiv:2606.06197v1 Announce Type: new Abstract: Question answering (QA) systems have achieved notable progress with the advent of large language models (LLMs). However, they still face challenges in a

Learning Predictive Visuomotor Coordination

ApplicationsDGX agent

arXiv:2503.23300v2 Announce Type: replace Abstract: Understanding and predicting human visuomotor coordination is crucial for applications in robotics, human-computer interaction, and assistive techno

Pitfalls of Evaluating Language Models with Open Benchmarks

SafetyDGX agent

arXiv:2507.00460v3 Announce Type: replace Abstract: Open Large Language Model (LLM) benchmarks, such as HELM and BIG-Bench, provide standardized and transparent evaluation protocols that support compa

Robust Scene Transfer for PointGoal Navigation via Privileged Sensor Guided Contrastive Learning

SafetyDGX agent

arXiv:2606.05506v1 Announce Type: new Abstract: We propose a sensor-guided adaptive contrastive learning framework for visual representation learning in PointGoal navigation. During training, privileg

Safe Embodied AI for Long-horizon Tasks: A Cross-layer Analysis of Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.05660v1 Announce Type: new Abstract: Embodied AI systems are increasingly expected to reason and act over extended horizons in physical environments. This growing capability brings safety t

TopoPult-SSL: Gland-Mask-Free Cross-Device Meibomian Gland Segmentation via Self-Distilled Weak Clinical Priors

Model ReleasesDGX agent

arXiv:2606.05347v1 Announce Type: new Abstract: Every new clinical imaging device creates a domain shift where dense gland masks are expensive yet cheap clinical signals -- eyelid outlines, Pult grade

Towards a Data Flywheel for Embodied Intelligence in Logistics

SafetyDGX agent

arXiv:2606.05960v1 Announce Type: new Abstract: Embodied intelligence is moving from laboratory demonstrations toward industrial deployment, with the logistics industry serving as a key application sc

Unsupervised Skill Discovery for Agentic Data Analysis

AgentsDGX agent

arXiv:2606.06416v1 Announce Type: cross Abstract: Inference-time skill augmentation provides a lightweight way to improve data-analytic agents by injecting reusable procedural knowledge without updati

Using street view images and visual LLMs to predict heritage values for governance support: Risks, ethics, and policy implications

SafetyDGX agent

arXiv:2601.06056v2 Announce Type: replace-cross Abstract: During 2025 and 2026, the Energy Performance of Buildings Directive is being implemented in the European Union member states, requiring all me

Would you still call this Dax? Novel Visual References in VLMs and Humans

Model ReleasesDGX agent

arXiv:2606.05409v1 Announce Type: cross Abstract: Vision-language models (VLMs), like human learners, are frequently exposed to new visual concepts, but how they map novel visual references to languag

4 Jun 2026

Affordance2Action: Task-Conditioned Scene-level Affordance Grounding for Real-Time Manipulation

Model ReleasesDGX agent

arXiv:2606.04172v1 Announce Type: new Abstract: Task-conditioned manipulation requires grounding instructions to task-relevant functional parts rather than object categories. This setting is scene-dep

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety

Model ReleasesDGX agent

arXiv:2606.04867v1 Announce Type: new Abstract: As AI companion platforms such as Replika and Character.AI rapidly grow, concerns about unsafe human-AI interactions have intensified. This study introd

An Open-Source Two-Stage Computer Vision Pipeline for Fine-Grained Vehicle Classification using Vision Transformers

Model ReleasesDGX agent

arXiv:2606.05149v1 Announce Type: new Abstract: Vehicle body type is a significant determinant of cyclist injury severity in overtaking crashes, yet automated tools for classifying vehicles into injur

Analysis-Driven Procedural Generation of an Engine Sound Dataset with Embedded Control Annotations

Model ReleasesDGX agent

arXiv:2603.07584v2 Announce Type: replace-cross Abstract: Computational engine sound modeling is central to the automotive audio industry, particularly for active sound design applications and virtual

Be Fair! Can Machine Learning Engineering Agents Adhere to Fairness Constraints?

SafetyDGX agent

arXiv:2606.04971v1 Announce Type: new Abstract: Machine learning engineering (MLE) agents promise to automate end-to-end ML pipeline development from raw data and natural language instructions, potent

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

Beyond Objective Equivalence: Constraint Injection for LLM-Based Optimization Modeling on Vehicle Routing Problems

Model ReleasesDGX agent

arXiv:2606.04816v1 Announce Type: new Abstract: Large language models (LLMs) increasingly translate natural-language optimization problems into executable solver code. Yet for constraint-dense operati

CADET: A Modular Platform for Evaluating Distributed Cooperative Autonomy in Connected Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2606.04072v1 Announce Type: cross Abstract: Deep learning models are increasingly central to autonomous vehicle (AV) pipelines, yet their integration has traditionally followed a monolithic desi

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

Model ReleasesDGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

DiffAero: A GPU-Accelerated Differentiable Simulation Framework for Efficient Quadrotor Policy Learning

SafetyDGX agent

arXiv:2509.10247v1 Announce Type: cross Abstract: This letter introduces DiffAero, a lightweight, GPU-accelerated, and fully differentiable simulation framework designed for efficient quadrotor contro

Dynamic Multi-Pair Trading Strategy in Cryptocurrency Markets with Deep Reinforcement Learning

SafetyDGX agent

arXiv:2606.04574v1 Announce Type: new Abstract: This study aims to determine whether the application of Deep Reinforcement Learning (DRL) as a specialized execution overlay can enhance pair trading in

FinTradeBench: A Financial Reasoning Benchmark for LLMs

Model ReleasesDGX agent

arXiv:2603.19225v3 Announce Type: replace-cross Abstract: Real-world financial decision-making is a challenging problem that requires reasoning over heterogeneous signals, including company fundamenta

Fog of Love: Engineering Virtuous Agent Behavior with Affinity-based Reinforcement Learning in a Game Environment

SafetyDGX agent

arXiv:2606.04750v1 Announce Type: new Abstract: Instilling virtuous behavior in artificial intelligence has seen increasing interest. One of the techniques proposed is known as affinity-based reinforc

Formal Semantics for Agentic Tool Protocols: A Process Calculus Approach

SafetyDGX agent

arXiv:2603.24747v2 Announce Type: replace Abstract: The emergence of large language model agents capable of invoking external tools has created urgent need for formal verification of agent protocols.

From Prompt to Process: a Process Taxonomy and Comparative Assessment of Frameworks Supporting AI Software Development Agents

AgentsDGX agent

arXiv:2606.04967v1 Announce Type: cross Abstract: AI tools for programming are no longer just autocomplete or chat assistants: they organize themselves as development frameworks, with process, roles,

here they are btw https://huggingface.co/datasets/clem/nanoclaw-traces

IndustryDGX agent

The nanoclaw-traces dataset on Hugging Face contains trace data related to nanoclaw, likely including execution logs, system calls, or behavioral patterns useful for analysis, debugging, or machine le

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-l…

Model ReleasesDGX agent

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-latency multilingual speech recognition. AI natives can now b

its been such fun befriending Pari and seeing him completely reinvent his company for the agentic era, WHILE having the most insanely stacke…

AgentsDGX agent

its been such fun befriending Pari and seeing him completely reinvent his company for the agentic era, WHILE having the most insanely stacked customer base I've ever seen in the hardest engineering do

← Previous
1…392393394395396…428
Next →