AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,911 results
10 Apr 2026

Self-Discovered Intention-aware Transformer for Multi-modal Vehicle Trajectory Prediction

AgentsDGX agent

arXiv:2604.07126v1 Announce Type: cross Abstract: Predicting vehicle trajectories plays an important role in autonomous driving and ITS applications. Although multiple deep learning algorithms are dev

Telescope: Learnable Hyperbolic Foveation for Ultra-Long-Range Object Detection

AgentsDGX agent

arXiv:2604.06332v1 Announce Type: cross Abstract: Autonomous highway driving, especially for long-haul heavy trucks, requires detecting objects at long ranges beyond 500 meters to satisfy braking dist

Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.08362v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) has illuminated the potential for a general-purpose user simulator. However, existing benchmarks remain co

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

Model ReleasesDGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

Uncertainty Estimation for Deep Reconstruction in Actuatic Disaster Scenarios with Autonomous Vehicles

AgentsDGX agent

arXiv:2604.06387v1 Announce Type: cross Abstract: Accurate reconstruction of environmental scalar fields from sparse onboard observations is essential for autonomous vehicles engaged in aquatic monito

Vercel CLI commands now scoped to local directory

ToolsDGX agent

Commands like `vc project ls` and `vc domains ls` now automatically use the scope of the linked local directory instead of defaulting to the global team. Previously, running these commands inside a...

9 Apr 2026

Cognition is live in Japan! 🇯🇵 We launched today with Tokyo Merge - a standing room only event with industry leaders across the country. W…

AgentsDGX agent

Cognition is live in Japan! 🇯🇵 We launched today with Tokyo Merge - a standing room only event with industry leaders across the country. We’re hiring, partnering, and building in Japan. If you’re buil

maybe some nuance 😄 I don’t think anyone is “lying” about how great Mythos will be —> but there’s expectation misalignment between the Test…

Model ReleasesDGX agent

maybe some nuance 😄 I don’t think anyone is “lying” about how great Mythos will be —> but there’s expectation misalignment between the Test Harness set up for Mythos and a belief it was given this cra

Sigma Automate emerges with $2.75M to tackle enterprise IT complexity with no-code automation

ApplicationsDGX agent

Sigma Automate Inc., a startup providing enterprises with no-code information technology automation, formally launched today with $2.75 million in funding to scale up its platform and expand its abili

Tesla puts a lot of effort into ensuring that our cars don’t run over animals

AgentsDGX agent

Tesla puts a lot of effort into ensuring that our cars don’t run over animals More than 350 million vertebrate animals are killed by human driven vehicles every year. In general, autonomous vehicles a

We have @TejasKumar_ at @aiDotEngineer teaching us about AI Harnesses 👀

ToolsDGX agent

Tejas Kumar ( a keynote speaker and developer with 20+ years of engineering experience across AI and web technologies ) presented on AI Harnesses at the AI Engineer Europe conference (April 8–10, 2...

Who’s building with AI…but stuck? We’re here to help you 🤝 Ask us anything! @Replit @samuel_spitz https://x.com/i/spaces/1RKZzjkoYRAKB

ToolsDGX agent

This X (Twitter) Space, hosted by @TheBestOfAdam and featuring Replit's Samuel Spitz, was a live Q&A session aimed at helping people who are building with AI but encountering obstacles or roadblock...

8 Apr 2026

Japanese railways don’t just run trains. They actually build cities and act as urban developers to capture the demand from transit. ~50% of …

AgentsDGX agent

Japanese railways don’t just run trains. They actually build cities and act as urban developers to capture the demand from transit. ~50% of revenue comes from owning the land around stations. Buy land

Read part 5 of our 6-part PM series: https://blog.replit.com/vibe-coding-decks-and-dashboards

ToolsDGX agent

Part 5 of Replit's 6-part PM series, 'Vibe Coding Decks and Dashboards,' covers how product managers can use Replit's agentic workflows to produce slide decks, dashboards, and launch assets without...

SWE-1.6 is lightning fast! Here's what 950 tok/s feels like - available in Windsurf today.

AgentsDGX agent

SWE-1.6 is lightning fast! Here's what 950 tok/s feels like - available in Windsurf today. Media We’re releasing SWE-1.6, our best model in both intelligence & model UX. SWE-1.6 matches our Preview mo

7 Apr 2026

GLM 5.1 is now LIVE in Atomic Chat SOTA for code & chat – now runs locally with TurboQuant Thanks to @zai_org for open-sourcing this frontie…

Model ReleasesDGX agent

GLM-5.1 is Z.ai's (zai-org) next-generation open-source flagship model for agentic engineering, achieving state-of-the-art performance on SWE-Bench Pro and significantly outperforming its predecess...

SWE-1.6 is free for everyone in Windsurf for the next 3 months at 200 tok/s. For paying users, we've partnered with Cerebras to serve the mo…

AgentsDGX agent

SWE-1.6 is free for everyone in Windsurf for the next 3 months at 200 tok/s. For paying users, we've partnered with Cerebras to serve the model at 950 tok/s. More technical details about this training

We’re releasing SWE-1.6, our best model in both intelligence & model UX. SWE-1.6 matches our Preview model on SWE-Bench Pro while dramatical…

AgentsDGX agent

We’re releasing SWE-1.6, our best model in both intelligence & model UX. SWE-1.6 matches our Preview model on SWE-Bench Pro while dramatically improving on various behavioral axes. It’s available toda

With SWE-1.6 we've made significant progress on 'intelligence per token'. We post-trained the model from scratch (same pre-trained model) wi…

AgentsDGX agent

With SWE-1.6 we've made significant progress on 'intelligence per token'. We post-trained the model from scratch (same pre-trained model) with a similar recipe as SWE-1.6 Preview. Our latest algorithm

14 Aug 2026

A Unifying Perspective on Causal World Models: From Observations to Representations to Structure

ResearchDGX agent

arXiv:2608.13456v1 Announce Type: new Abstract: World Models (WM) are increasingly seen as a foundation for intelligent agents that can predict, plan, and act beyond their training distribution. In th

AirForesight: Current-to-Future Spatial Map Imagination with Cross-Space Planning Consistency for UAV-VLN

ResearchDGX agent

arXiv:2608.12835v1 Announce Type: new Abstract: Unmanned Aerial Vehicle Vision-Language Navigation (UAV-VLN) requires agents to follow language instructions, infer spatial structure from sparse multi-

vToken: Token-Level Virtualization for Reclaimable KV Caches

SafetyDGX agent

arXiv:2608.13263v1 Announce Type: new Abstract: Large language model serving faces a critical memory bottleneck: the KV cache grows with sequence length and batch size. PagedAttention uses fixed-size

13 Aug 2026

A weird experiment I've been trying the last few weeks is having Claude take over day-to-day maintenance of our apps. Seeing early signs of …

Model ReleasesDGX agent

A weird experiment I've been trying the last few weeks is having Claude take over day-to-day maintenance of our apps. Seeing early signs of life that this might be possible. The setup is straightforwa

Better Slots, Better Worlds: Representation Quality & Robustness in Object-Centric World Models

SafetyDGX agent

arXiv:2608.12078v1 Announce Type: cross Abstract: Learning world models from offline trajectories enables agents to accomplish different tasks through planning. Object-centric (OC) representations, wh

Deepseek Harness is Up!

Model ReleasesDGX agent

DeepSeek Harness (dsh) is an open-source agent harness developed by DeepSeek AI. It uses an architecture where everything is a plugin, and is powered by Cordis, whose design is described in A Programm

DreamFly: Causal Memory and Receding-Horizon Diffusion Planning for Aerial Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2608.12308v1 Announce Type: cross Abstract: Aerial vision-language navigation (VLN) requires an embodied agent to integrate visual evidence over time, plan future actions, and determine when it

Learning to Persuade Exposes How Easily LLMs Abandon Correct Beliefs

Model ReleasesDGX agent

arXiv:2608.11624v1 Announce Type: cross Abstract: Persuasion is a core dynamic of natural language communication, shaping how large language models (LLMs) update beliefs, resolve disagreements, and re

Map-Det3D: Metric Feed-Forward 3D Reconstruction Prior for Multi-view 3D Object Detection from Streaming Inputs

Local AiDGX agent

arXiv:2608.12179v1 Announce Type: new Abstract: Metric 3D object detection is a core capability for embodied agents, yet most reliable systems lean on depth sensors, trading away cost, power, and inte

MaSRead: Content-Addressed Reading of Replicated Latent Stores

ResearchDGX agent

arXiv:2608.11218v1 Announce Type: new Abstract: Independent agents that reason in latent space can share computed state as key-value cache fragments rather than text. Merged by a conflict-free replica

Program Semantic Inequivalence Game with Large Language Models

Model ReleasesDGX agent

arXiv:2505.03818v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) can achieve strong performance on everyday coding tasks, but they can fail on complex tasks that require non-triv

The Off-Support Barrier: Why Semantic Safety Constraints Are Not Learning-Problem Invariants, and What Follows for Prior Design, Containment, and Verification

Local AiDGX agent

arXiv:2608.11243v1 Announce Type: new Abstract: We argue that a single structural fact organizes a wide range of phenomena in contemporary AI safety: a semantic safety constraint (e.g., the agent does

12 Aug 2026

18 two-word AI prompts I'm kind of obsessed with: 1) now what - great for when you've wrapped up a project or big push and you still have en…

Model ReleasesDGX agent

18 two-word AI prompts I'm kind of obsessed with: 1) now what - great for when you've wrapped up a project or big push and you still have energy and want AI to give you more 2) plz fix - usually accom

Azure Content Understanding GPT-5 Series Guide: Model Selection, Grounding Improvements, and Confidence Enhancements

Model ReleasesDGX agent

Enterprise content is no longer just something people consume. As organizations increasingly rely on AI to extract and act on information from documents, images, audio, and video, Azure Content Unders

Easy3D-Labels: Supervising Semantic Occupancy Estimation with 3D Pseudo-Labels for Automotive Perception

Local AiDGX agent

arXiv:2509.26087v5 Announce Type: replace Abstract: In perception for automated vehicles, safety is critical not only for the driver but also for other agents in the scene, particularly vulnerable roa

Every Token Counts: Exact Likert-Scale Distributions for Measuring LLM Attitudes and Biases

SafetyDGX agent

arXiv:2608.10503v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as autonomous agents, accurately evaluating their latent values and biases is critical. The NL

EvoMem: Memory-Augmented Evolution for Code Optimization

HardwareDGX agent

arXiv:2608.10795v1 Announce Type: new Abstract: Successful mutation strategies in evolutionary code search may contain reusable knowledge that is useful beyond a single run, and in some cases may tran

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration.

Model ReleasesDGX agent

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration. SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, j

Grok 4.6 is objectively #1 when considering intelligence, speed & cost

Model ReleasesDGX agent

Grok 4.6 is objectively #1 when considering intelligence, speed & cost SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, joining the frontier in line with GPT-5.6 Sol, with

Grok Bot

Model ReleasesDGX agent

Grok Bot Here's my Grok Bot team: - Webby: Web designer - Shotry: Short-form content creator - Writey: Article/Newsletter writer - Claude Code: Grok agent that specializes in CC - Codex: Same as the a

ImpactHO: Importance-Aware KV Cache Transfer for Multi-User Edge LLM Handover

Model ReleasesDGX agent

arXiv:2608.10545v1 Announce Type: cross Abstract: Edge LLMs must preserve inference continuity when a user hands over between edge nodes, requiring key-value (KV) cache transfer to the target node. Ho

Most biomedical publications show signs of LLM-assisted writing

SafetyDGX agent

arXiv:2608.10715v1 Announce Type: cross Abstract: Over the past several years, LLM-powered chatbots and agents have become widely used as a tool for academic writing. LLM-assisted writing can be valua

New Muse-Glimmer-30B SoTA Quants - hopefully a new lineup :)

Model ReleasesDGX agent

Hey Folks, I've been making quants for a while - recently I took a short break to get into hardcore research (submitted my first EMNLP paper during it!). Along the way, I built up a little arsenal of

Partially Observable Learning for Multi-Platform Dispatch Optimization

SafetyDGX agent

arXiv:2608.10897v1 Announce Type: new Abstract: Instant delivery platforms have become a critical component of urban logistics, increasingly relying on crowdsourced couriers to fulfill highly dynamic

RLMOpt: Adaptive Prompt Optimization via Recursive Language Models

Model ReleasesDGX agent

arXiv:2608.10471v1 Announce Type: new Abstract: Prompt optimizers automate the search for prompts that improve language-model performance, but existing methods rely on a predefined optimization proced

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning

SafetyDGX agent

arXiv:2608.10513v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) remain vulnerable to jailbreak attacks that exploit visual inputs to bypass safety alignment inherited from their

Sheaf-Based Federated Representation Learning

Local AiDGX agent

arXiv:2608.10016v1 Announce Type: cross Abstract: Heterogeneous federated systems require agents to learn and exchange informative representations despite differences in data distributions, sensing mo

Situation Graph Prediction for User Perspective Modeling

Model ReleasesDGX agent

arXiv:2602.13319v2 Announce Type: replace Abstract: Perspective-aware AI requires modeling evolving internal states---goals, emotions, contexts---not merely preferences. Progress is limited by a data

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs

Model ReleasesDGX agent

arXiv:2608.10042v1 Announce Type: cross Abstract: Tool-use LLMs are increasingly asked to act on users' behalf, but existing benchmarks usually focus on profile recall, style imitation, generic tool u

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: …

Model ReleasesDGX agent

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: 🧠 LLMs & Reasoning → DeepSeek-V4-Flash-0731 (my king 👑): 304B

11 Aug 2026

Action- and Language-Conditioned Video Assessment for Embodied Control

SafetyDGX agent

arXiv:2608.08273v1 Announce Type: cross Abstract: Vision-based embodied agents executing multi-step natural language instructions require feedback mechanisms that assess task progress over complete tr

Avalon-ToM-Bench: Evaluating Fine-Grained Theory of Mind via Asymmetric Game Mechanics

Model ReleasesDGX agent

arXiv:2608.09638v1 Announce Type: new Abstract: Theory of Mind (ToM) is essential for agent interactions, yet existing evaluations either rely on static scenarios that oversimplify mental-state reason

Carnot: Interpretable, Interactive, and Optimized Execution of Deep Research Queries

ApplicationsDGX agent

arXiv:2608.09532v1 Announce Type: cross Abstract: Enterprises increasingly seek to query data lakes using natural language via AI-driven tools like semantic operators or deep research agents. However,

CausalNav: Reliability-Certified Causal World Models for Control under Physical-Parameter Shift

Model ReleasesDGX agent

arXiv:2608.07809v1 Announce Type: new Abstract: A world model is only useful for physical AI if it changes what the agent does, and only safe if it declines to do so when it is wrong. We study both ha

Concept-Guided Spatial Regularization for World Models in Atari Pong

SafetyDGX agent

arXiv:2607.15142v2 Announce Type: replace Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, leaving their standalone reliability understu

Curriculum Generation under Structured Parametric Environments for Robust Navigation Policies

Model ReleasesDGX agent

arXiv:2608.08545v1 Announce Type: cross Abstract: Robust navigation policies for autonomous agents must generalize across continuously varying environmental conditions such as turn rates, obstacles, f

DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO…

Model ReleasesDGX agent

DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO, then deploy the fine-tuned model on Together AI for produc

Diminishing Returns of Intelligence: The Non-Linear Relationship Between LLM Scale and User Perception in Short-Duration Open-Ended Social Human-Robot Interactions

Model ReleasesDGX agent

arXiv:2608.08320v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to drive embodied social agents, yet it remains unclear whether larger models improve user perception

DSLE: A Learning Environment for Dark Souls Boss Encounters

SafetyDGX agent

arXiv:2608.09902v1 Announce Type: new Abstract: We introduce the Dark Souls Learning Environment (DSLE), a containerized platform that presents all 22 boss encounters of Dark Souls: Remastered as game

From Semantic Grounding to Decision Optimization: A Unified Framework for Long-Horizon UAV Vision-Language Navigation

SafetyDGX agent

arXiv:2608.09564v1 Announce Type: cross Abstract: UAV vision-language navigation (UAV-VLN) focuses on enabling an aerial agent to follow natural-language instructions in open 3D environments from egoc

GraphThink: Graph-Enhanced LLM Thinking for Long-Horizon Embodied Task Planning

Model ReleasesDGX agent

arXiv:2608.07905v1 Announce Type: new Abstract: Embodied agents using LLM-based planners often struggle with physical hallucinations, poor generalization to long-horizon tasks, and lack of environment

← Previous
1…224225226227228…299
Next →