AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
Model Releases

TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications

DGX agent

arXiv:2605.09929v1 Announce Type: new Abstract: Deploying large language models in telecommunications requires more than task accuracy. In realistic workflows, a model may inherit partially completed

model-releasesarxiv-cs-lg
12 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other techn…

DGX agent

There will be no AI jobpocalypse. The story that AI will lead to massive unemployment is stoking unnecessary fear. AI — like any other technology — does affect jobs, but telling overblown stories of l

safetyandrew-ng--x
12 May 2026
Safety

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

DGX agent

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

safetyyohei-nakajima--x
12 May 2026
Safety

Towards Customized Multimodal Role-Play

DGX agent

arXiv:2605.08129v1 Announce Type: new Abstract: Unified multimodal understanding and generation models enable richer human-AI interaction. Yet jointly customizing a character's persona, dialogue style

safetyarxiv-cs-lg
12 May 2026
Safety

Training-Free Cultural Alignment of Large Language Models via Persona Disagreement

DGX agent

arXiv:2605.10843v1 Announce Type: cross Abstract: Large language models increasingly mediate decisions that turn on moral judgement, yet a growing body of evidence shows that their implicit preference

safetyarxiv-cs-ai
12 May 2026
Model Releases

UserGPT Technical Report

DGX agent

arXiv:2605.08766v1 Announce Type: cross Abstract: Personalized user understanding from large-scale digital traces remains a fundamental challenge. Traditional user profiling methods rely on discrimina

model-releasesarxiv-cs-cl
12 May 2026
Safety

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning

DGX agent

arXiv:2605.10172v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved remarkable success in general perception, yet complex multi-step visual reasoning remains a per

safetyarxiv-cs-cl
12 May 2026
Industry

We just crossed 1,000,000 public datasets on Hugging Face! That's petabytes of data available that millions of AI builders are downloading, …

DGX agent

We just crossed 1,000,000 public datasets on Hugging Face! That's petabytes of data available that millions of AI builders are downloading, analyzing, and training AI models on every day! What's inter

industryclem-delangue--x
12 May 2026
Safety

When a Robot is More Capable than a Human: Learning from Constrained Demonstrators

DGX agent

arXiv:2510.09096v3 Announce Type: replace-cross Abstract: Learning from demonstrations enables experts to teach robots complex tasks using interfaces such as kinesthetic teaching, joystick control, an

safetyarxiv-cs-ai
12 May 2026
Model Releases

2.5-D Decomposition for LLM-Based Spatial Construction

DGX agent

arXiv:2605.07066v1 Announce Type: new Abstract: Autonomous systems that build structures from natural-language instructions need reliable spatial reasoning, yet large language models (LLMs) make syste

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning

DGX agent

arXiv:2502.07143v3 Announce Type: replace Abstract: The severe shortage of medical doctors limits access to timely and reliable healthcare, leaving millions underserved. Large language models (LLMs) o

model-releasesarxiv-cs-cl
11 May 2026
Hardware

CktFormalizer: Autoformalization of Natural Language into Circuit Representations

DGX agent

arXiv:2605.07782v1 Announce Type: new Abstract: LLMs can generate hardware descriptions from natural language specifications, but the resulting Verilog often contains width mismatches, combinational l

hardwarearxiv-cs-cl
11 May 2026
Model Releases

Cloud Storage Rapid: Turbocharged object storage for AI and analytics

DGX agent

At Google Cloud Next ’26 we announced Cloud Storage Rapid, a family of object storage capabilities for data-intensive workloads like AI and analytics. Out of the gate, Cloud Storage Rapid consists of

model-releasesgoogle-cloud-ai
11 May 2026
Model Releases

DeepSeek V4 Flash is ~90% cheaper than GPT 5.4 Mini and ~70% cheaper than Gemini 3.1 Flash Lite For devs pushing ~500M tok/month, this is th…

DGX agent

DeepSeek V4 Flash is ~90% cheaper than GPT 5.4 Mini and ~70% cheaper than Gemini 3.1 Flash Lite For devs pushing ~500M tok/month, this is the difference between: GPT 5.4 Mini: ~394/mo Gemini 3.1 Flash

model-releasesharrison-chase--x
11 May 2026
Model Releases

Do Joint Audio-Video Generation Models Understand Physics?

DGX agent

arXiv:2605.07061v1 Announce Type: cross Abstract: Joint audio-video generation models are rapidly approaching professional production quality, raising a central question: do they understand audio-visu

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

EgoPro-Bench: Benchmarking Personalized Proactive Interaction in Egocentric Video Streams

DGX agent

arXiv:2605.07299v1 Announce Type: cross Abstract: Existing Multimodal Large Language Models (MLLMs) remain primarily reactive, failing to continuously perceive environments or proactively assist users

model-releasesarxiv-cs-ai
11 May 2026
Safety

Emergent social transmission of model-based representations without inference

DGX agent

arXiv:2604.05777v2 Announce Type: replace Abstract: How do people acquire rich, flexible knowledge about their environment from others despite limited cognitive capacity? Humans are often thought to r

safetyarxiv-cs-ai
11 May 2026
Model Releases

How Far Are VLMs from Privacy Awareness in the Physical World? An Empirical Study

DGX agent

arXiv:2605.05340v2 Announce Type: replace-cross Abstract: As Vision-Language Models (VLMs) are increasingly deployed as autonomous cognitive cores for embodied assistants, evaluating their privacy awa

model-releasesarxiv-cs-ai
11 May 2026
Safety

InvThink: Premortem Reasoning for Safer Language Models

DGX agent

arXiv:2510.01569v3 Announce Type: replace Abstract: We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before gener

safetyarxiv-cs-ai
11 May 2026
Model Releases

Meet the latest Database Center, now with Gemini-powered fleet intelligence

DGX agent

Managing a modern database fleet is both a scale and cognitive problem. As database estates grow, the effort required to monitor, troubleshoot, and optimize them often outpaces teams’ capacity, who fi

model-releasesgoogle-cloud-ai
11 May 2026
Model Releases

MiniAppBench: Evaluating the Shift from Text to Interactive HTML Responses in LLM-Powered Assistants

DGX agent

arXiv:2603.09652v3 Announce Type: replace Abstract: With the rapid advancement of Large Language Models (LLMs) in code generation, human-AI interaction is evolving from static text responses to dynami

model-releasesarxiv-cs-ai
11 May 2026
Safety

Operating Within the Operational Design Domain: Zero-Shot Perception with Vision-Language Models

DGX agent

arXiv:2605.07649v1 Announce Type: cross Abstract: Over the last few years, research on autonomous systems has matured to such a degree that the field is increasingly well-positioned to translate resea

safetyarxiv-cs-ai
11 May 2026
Model Releases

RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation

DGX agent

arXiv:2605.07334v1 Announce Type: new Abstract: Video Reasoning Segmentation (VRS) aims to segment target objects in videos based on implicit instructions that convey human intent and temporal logic.

model-releasesarxiv-cs-cv
11 May 2026
Research

Revisiting Adam for Streaming Reinforcement Learning

DGX agent

arXiv:2605.06764v1 Announce Type: cross Abstract: Learning from a sequence of interactions, as soon as observations are perceived and acted upon, without explicitly storing them, holds the promise of

researcharxiv-cs-ai
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

DGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

model-releasesarxiv-cs-cl
11 May 2026
Research

Semantic State Abstraction Interfaces for LLM-Augmented Portfolio Decisions: Multi-Axis News Decomposition and RL Diagnostics

DGX agent

arXiv:2605.06730v1 Announce Type: new Abstract: We introduce Semantic State Abstraction Interfaces (SSAI): a methodological template for mapping sparse unstructured text into K auditable, named coordi

researcharxiv-cs-lg
11 May 2026
Hardware

🧵 Slime: The Most Elegant & Comfortable RL Training Framework Ever A deep dive into why Slime redefines LLM RL training with clean architec…

DGX agent

🧵 Slime: The Most Elegant & Comfortable RL Training Framework Ever A deep dive into why Slime redefines LLM RL training with clean architecture & production-grade engineering ✨ Insights from Zhihu con

hardwarezhipu-ai--x
11 May 2026
Model Releases

The Position Curse: LLMs Struggle to Locate the Last Few Items in a List

DGX agent

arXiv:2605.07127v1 Announce Type: cross Abstract: Modern large language models (LLMs) can find a needle in a haystack (locating a single relevant fact buried among hundreds of thousands of irrelevant

model-releasesarxiv-cs-cl
11 May 2026
Local Ai

Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles

DGX agent

arXiv:2512.03454v4 Announce Type: replace-cross Abstract: Interpreting natural-language commands to localize target objects is critical for autonomous driving (AD). Existing visual grounding (VG) meth

local-aiarxiv-cs-ai
11 May 2026
Tools

“we gonna yolo our way into running the biggest conf in town.” and somehow… we actually did it with @aiDotEngineer singapore a year ago, @ag…

DGX agent

“we gonna yolo our way into running the biggest conf in town.” and somehow… we actually did it with @aiDotEngineer singapore a year ago, @agrimsingh @unprofeshme and i joked about the idea of running

toolsswyx--x
11 May 2026
Tutorials

Your AI Use Is Breaking My Brain

DGX agent

Your AI Use Is Breaking My Brain Excellent, angry piece by Jason Koebler on how AI writing online is becoming impossible to avoid, filtering it is mentally exhausting and it's even starting to distort

tutorialssimon-willison
11 May 2026
Applications

A look at Janitor AI, a romantic fantasy roleplay chatbot site run by three men that claims 2.5M DAUs and 15M total users, with 70% to 80% identifying as women (Anna Tong/Forbes)

DGX agent

Anna Tong / Forbes: A look at Janitor AI, a romantic fantasy roleplay chatbot site run by three men that claims 2.5M DAUs and 15M total users, with 70% to 80% identifying as women — Despite its name,

applicationstechmeme
10 May 2026
Model Releases

100% agree on the Context Hub. Developers constantly tweak their approaches to manage context in their prompts with each new model and tool …

DGX agent

100% agree on the Context Hub. Developers constantly tweak their approaches to manage context in their prompts with each new model and tool suite release. I would even say that the core problem we fac

model-releasesharrison-chase--x
9 May 2026
Model Releases

ERNIE 5.1 is here 🚀 ERNIE 5.1 significantly reduces pretraining cost while compressing total parameters to ~1/3 and activated parameters to…

DGX agent

ERNIE 5.1 is here 🚀 ERNIE 5.1 significantly reduces pretraining cost while compressing total parameters to ~1/3 and activated parameters to ~1/2 — using only ~6% of the pretraining cost compared to mo

model-releasesjeremy-howard--x
9 May 2026
Model Releases

I recently installed Hermes after raw-dogging Claude Code for a year or so. Have tried it at work, didn't see the need to implement it perso…

DGX agent

I recently installed Hermes after raw-dogging Claude Code for a year or so. Have tried it at work, didn't see the need to implement it personally. Always had the usual fears about 'privacy' and what n

model-releasesnous-research--x
9 May 2026
Hardware

Improving Bash Generation in Small Language Models with Grammar-Constrained Decoding

DGX agent

Grammar-constrained decoding modifies language model generation by applying grammar constraints at each step to block structurally invalid tokens , ensuring syntactically correct Bash command generati

hardwarenvidia-developer
8 May 2026
Safety

Remember that MIT study that showed that the ROI for generative AI wasn’t really there for most businesses? Or any of the six or seven studi…

DGX agent

Remember that MIT study that showed that the ROI for generative AI wasn’t really there for most businesses? Or any of the six or seven studies from other teams that followed, showing basically the sam

safetygary-marcus--x
8 May 2026
Tutorials

Thank you to @robertwiblin for inviting me on the @80000Hours podcast to discuss the research progress we’re making at @LawZero_ to create s…

DGX agent

Thank you to @robertwiblin for inviting me on the @80000Hours podcast to discuss the research progress we’re making at @LawZero_ to create safe-by-design AI systems. Our current approach, Scientist AI

tutorialsyoshua-bengio--x
8 May 2026
Industry

Which Macs are suffering from shortages—and where are things getting worse?

DGX agent

Apple CEO Tim Cook acknowledged in Q2 2026 earnings that high-demand Mac mini and Mac Studio configurations are severely constrained and may take several months to reach supply-demand balance. Apple h

industryars-technica
8 May 2026
Model Releases

AsymmetryZero: A Framework for Operationalizing Human Expert Preferences as Semantic Evals

DGX agent

arXiv:2605.04083v1 Announce Type: new Abstract: Much of the focus in RL today is on evaluation design: building meaningful evals that serve simultaneously as benchmarks and as well-defined reward sign

model-releasesarxiv-cs-lg
7 May 2026
Applications

Congrats on the launch! Filesystems are all you need (?) There wasn't a huge demand for 'managed RAG' services in 2023, but it's possible th…

DGX agent

Congrats on the launch! Filesystems are all you need (?) There wasn't a huge demand for 'managed RAG' services in 2023, but it's possible the infra and market was just not mature enough. Maybe filesys

applicationsjerry-liu--x
7 May 2026
Local Ai

Contextual Multi-Objective Optimization: Rethinking Objectives in Frontier AI Systems

DGX agent

arXiv:2605.03900v1 Announce Type: new Abstract: Frontier AI systems perform best in settings with clear, stable, and verifiable objectives, such as code generation, mathematical reasoning, games, and

local-aiarxiv-cs-ai
7 May 2026
Research

Hacker News → LLM Artifact I built the most personalized HN feed. It only tracks topics I do research around based on memory and LLM wiki. N…

DGX agent

Hacker News → LLM Artifact I built the most personalized HN feed. It only tracks topics I do research around based on memory and LLM wiki. No point in storing bookmarks. With a few automations, rules,

researchdair-ai--x
7 May 2026
Safety

I really enjoyed chatting with @mattturck, was a great discussion.

DGX agent

I really enjoyed chatting with @mattturck, was a great discussion. Deeply thoughtful conversation with @zicokolter, board member at @OpenAI and head of the machine learning department at @CarnegieMell

safetyjeremy-howard--x
7 May 2026
Safety

InterFuserDVS: Event-Enhanced Sensor Fusion for Safe RL-Based Decision Making

DGX agent

arXiv:2605.04355v1 Announce Type: new Abstract: Autonomous driving systems rely heavily on robust sensor fusion to perceive complex envi- ronments. Traditional setups using RGB cameras and LiDAR often

safetyarxiv-cs-cv
7 May 2026
Model Releases

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning

DGX agent

arXiv:2602.04129v2 Announce Type: replace Abstract: Heterogeneous multi-robot systems are increasingly used in long-horizon missions requiring coordinated planning across diverse capabilities. However

model-releasesarxiv-cs-ro
7 May 2026
Local Ai

Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation

DGX agent

arXiv:2512.23864v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown remarkable generalization by mapping web-scale knowledge to robotic control, yet they remain bl

local-aiarxiv-cs-cv
7 May 2026
← Previous
1…350351352353354…367
Next →