AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
15 Apr 2026

Cursor can now respond by creating interactive canvases to visually represent information. Ask it to generate dashboards and custom interfac…

ToolsDGX agent

Cursor has introduced a new feature that allows its AI to respond by generating interactive canvases, enabling users to visually represent information beyond standard text responses. Users can prompt

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction

Model ReleasesDGX agent

arXiv:2603.07083v2 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) agents operating in high-dimensional observation spaces, such as Dreamer, rely on learning abstract repres

From Kinematics to Dynamics: Learning to Refine Hybrid Plans for Physically Feasible Execution

Research
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.12474v1 Announce Type: cross Abstract: In many robotic tasks, agents must traverse a sequence of spatial regions to complete a mission. Such problems are inherently mixed discrete-continuou

Hosting Live session for sub 10ms retrieval by Moss (YC backed) [N]

ResearchDGX agent

Moss is a YC-backed high-performance runtime for real-time semantic search that delivers sub-10ms lookups, instant index updates, and zero infrastructure overhead, running where the agent lives — clou

If it can’t parse, it can’t perform. ❌ Reliable document understanding is fundamental to enterprise-grade #AgenticAutomation. We're excited …

Model ReleasesDGX agent

If it can’t parse, it can’t perform. ❌ Reliable document understanding is fundamental to enterprise-grade #AgenticAutomation. We're excited to see the release of ParseBench, a new open-source benchmar

Memory as Metabolism: A Design for Companion Knowledge Systems

Model ReleasesDGX agent

arXiv:2604.12034v1 Announce Type: new Abstract: Retrieval-Augmented Generation remains the dominant pattern for giving LLMs persistent memory, but a visible cluster of personal wiki-style memory archi

ProbeLogits: Kernel-Level LLM Inference Primitives for AI-Native Operating Systems

Model ReleasesDGX agent

arXiv:2604.11943v1 Announce Type: cross Abstract: An OS kernel that runs LLM inference internally can read logit distributions before any text is generated -- and act on them as a governance primitive

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation

SafetyDGX agent

arXiv:2511.17097v2 Announce Type: replace Abstract: Vision-Language Navigation requires agents to act coherently over long horizons by understanding not only local visual context but also how far they

proud of this launch from a strong team. dedicated hardware for high throughput, low cost search building for the new era of recommenders an…

ApplicationsDGX agent

proud of this launch from a strong team. dedicated hardware for high throughput, low cost search building for the new era of recommenders and agents Dedicated Read Nodes are now generally available. P

Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.12371v1 Announce Type: new Abstract: We study typographic prompt injection attacks on vision-language models (VLMs), where adversarial text is rendered as images to bypass safety mechanisms

Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss

Model ReleasesDGX agent

arXiv:2604.12911v1 Announce Type: cross Abstract: Multilingual benchmarks guide the development of frontier models. Yet multilingual evaluations reported by frontier models are structured similar to p

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

SafetyDGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling

ResearchDGX agent

arXiv:2508.04282v3 Announce Type: replace Abstract: Recent benchmarks for memory-augmented reinforcement learning (RL) have introduced partially observable Markov decision process (POMDP) environments

TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs

SafetyDGX agent

arXiv:2604.12232v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed across diverse domains, yet their vulnerability to jailbreak attacks, where adversarial inputs

This sub is mostly screenshots of ChatGPT being wrong. Meanwhile people on RunLobster (OpenClaw) are quietly running real businesses on AI.

IndustryDGX agent

This Reddit post contrasts the r/ChatGPT community's tendency to focus on AI failures and viral screenshots with a more pragmatic user base building real workflows on RunLobster, a managed cloud platf

14 Apr 2026

8 AI and data trends shaping financial services in 2026

IndustryDGX agent

Databricks outlines eight key AI and data trends expected to transform financial services in 2026, likely covering advancements in generative AI, real-time data processing, machine learning for risk m

AffordGen: Generating Diverse Demonstrations for Generalizable Object Manipulation with Afford Correspondence

SafetyDGX agent

arXiv:2604.10579v1 Announce Type: cross Abstract: Despite the recent success of modern imitation learning methods in robot manipulation, their performance is often constrained by geometric variations

Big lab leaks

IndustryDGX agent

AI development is shifting toward fully integrated, agent-enabled applications, with major platforms like Anthropic's Claude and OpenAI

CoSToM:Causal-oriented Steering for Intrinsic Theory-of-Mind Alignment in Large Language Models

SafetyDGX agent

arXiv:2604.10031v1 Announce Type: cross Abstract: Theory of Mind (ToM), the ability to attribute mental states to others, is a hallmark of social intelligence. While large language models (LLMs) demon

Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations

SafetyDGX agent

arXiv:2604.11322v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive capabilities in utilizing external tools. In practice, however, LLMs are often exposed to to

ExecTune: Effective Steering of Black-Box LLMs with Guide Models

Model ReleasesDGX agent

arXiv:2604.09741v1 Announce Type: cross Abstract: For large language models deployed through black-box APIs, recurring inference costs often exceed one-time training costs. This motivates composed age

FlashMem: Distilling Intrinsic Latent Memory via Computation Reuse

Model ReleasesDGX agent

arXiv:2601.05505v2 Announce Type: replace Abstract: The stateless architecture of Large Language Models inherently lacks the mechanism to preserve dynamic context, compelling agents to redundantly rep

Had a fantastic time at the @aiDotEngineer Europe conference in London last week. There was a lot of good coverage of people's emerging appr…

ToolsDGX agent

Had a fantastic time at the @aiDotEngineer Europe conference in London last week. There was a lot of good coverage of people's emerging approaches for working with AI coding agents. A couple of the st

How some mathematicians are exploring ways to incorporate LLM models into their research without losing direct experience with mathematical understanding (Konstantin Kakaes/Quanta Magazine)

IndustryDGX agent

Konstantin Kakaes / Quanta Magazine: How some mathematicians are exploring ways to incorporate LLM models into their research without losing direct experience with mathematical understanding — Those c

However, most alignment research is not very crisp and requires research taste when evaluating. This is why we chose to point the AAR at thi…

SafetyDGX agent

However, most alignment research is not very crisp and requires research taste when evaluating. This is why we chose to point the AAR at this scalable oversight problem! Progress would let AARs work o

Learning to Focus: CSI-Free Hierarchical MARL for Reconfigurable Reflectors

SafetyDGX agent

arXiv:2604.05165v2 Announce Type: replace Abstract: Reconfigurable Intelligent Surfaces (RIS) has a potential to engineer smart radio environments for next-generation millimeter-wave (mmWave) networks

Like a Hammer, It Can Build, It Can Break: Large Language Model Uses, Perceptions, and Adoption in Cybersecurity Operations on Reddit

SafetyDGX agent

arXiv:2604.09998v1 Announce Type: cross Abstract: Large language models (LLMs) have recently emerged as promising tools for augmenting Security Operations Center (SOC) workflows, with vendors increasi

Panoptic Pairwise Distortion Graph

Model ReleasesDGX agent

arXiv:2604.11004v1 Announce Type: cross Abstract: In this work, we introduce a new perspective on comparative image assessment by representing an image pair as a structured composition of its regions.

ParseBench is the most comprehensive OCR benchmark for real-world enterprise documents: financial filings, contracts, insurance documents, a…

Model ReleasesDGX agent

ParseBench is the most comprehensive OCR benchmark for real-world enterprise documents: financial filings, contracts, insurance documents, and more. We evaluate across 5 dimensions that are present am

Real-Time Voicemail Detection in Telephony Audio Using Temporal Speech Activity Features

HardwareDGX agent

arXiv:2604.09675v1 Announce Type: cross Abstract: Outbound AI calling systems must distinguish voicemail greetings from live human answers in real time to avoid wasted agent interactions and dropped c

Retrieval Is Not Enough: Why Organizational AI Needs Epistemic Infrastructure

ResearchDGX agent

arXiv:2604.11759v1 Announce Type: new Abstract: Organizational knowledge used by AI agents typically lacks epistemic structure: retrieval systems surface semantically relevant content without distingu

Risk Awareness Injection: Calibrating Vision-Language Models for Safety without Compromising Utility

SafetyDGX agent

arXiv:2602.03402v3 Announce Type: replace Abstract: Vision language models (VLMs) extend the reasoning capabilities of large language models (LLMs) to cross-modal settings, yet remain highly vulnerabl

Scaling unstructured enterprise knowledge with BigQuery Graph, and Kineviz GraphXR

Model ReleasesDGX agent

Over 80% of enterprise data lives in unstructured form — PDFs, emails, reports, regulatory filings. Most of the time, such sources contain critical business information, yet they remain difficult to a

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

Model ReleasesDGX agent

arXiv:2505.17012v3 Announce Type: replace-cross Abstract: Existing evaluations of multimodal large language models (MLLMs) on spatial intelligence are typically fragmented and limited in scope. In thi

Start building a Sentry automation with a template from our marketplace: http://cursor.com/marketplace/automations/investigate-sentry-issues

ToolsDGX agent

Cursor offers a pre-built automation template in its marketplace that helps developers investigate Sentry issues more efficiently. The template, accessible at cursor.com/marketplace/automations/invest

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

Model ReleasesDGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

Taking a Pulse on How Generative AI is Reshaping the Software Engineering Research Landscape

SafetyDGX agent

arXiv:2604.11184v1 Announce Type: cross Abstract: Context: Software engineering (SE) researchers increasingly study Generative AI (GenAI) while also incorporating it into their own research practices.

The user interface of the future is your voice: Inside 8×8’s AI Studio

TutorialsDGX agent

For years, the promise of “no-code” artificial intelligence has felt a bit like a “some assembly required” IKEA desk — sure, you aren’t sawing the wood yourself, but you’re trying to decipher how to a

Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models

Model ReleasesDGX agent

arXiv:2604.10733v1 Announce Type: cross Abstract: Large language models increasingly serve as conversational agents that adopt personas and role-play characters at user request. This capability, while

13 Apr 2026

Adaptive Rigor in AI System Evaluation using Temperature-Controlled Verdict Aggregation via Generalized Power Mean

Model ReleasesDGX agent

arXiv:2604.08595v1 Announce Type: cross Abstract: Existing evaluation methods for LLM-based AI systems, such as LLM-as-a-Judge, verdict systems, and NLI, do not always align well with human assessment

Benchmarked @DJLougen ’s Ornstein-27B-v2 Q6_K on my RTX 3090 using hermes-bench, my new open-source benchmarking UI for local LLMs and Herme…

Model ReleasesDGX agent

Benchmarked @DJLougen ’s Ornstein-27B-v2 Q6_K on my RTX 3090 using hermes-bench, my new open-source benchmarking UI for local LLMs and Hermes agents. Ornstein is a Qwen 3.5 27B fine-tune trained on re

CaRLi-V: Camera-RADAR-LiDAR Point-Wise 3D Velocity Estimation

ResearchDGX agent

arXiv:2511.01383v2 Announce Type: replace Abstract: Accurate point-wise velocity estimation in 3D is crucial for robot interaction with non-rigid dynamic agents, enabling robust performance in path pl

China has erased the US lead in AI, Stanford HAI’s 2026 AI index reveals

Model ReleasesDGX agent

Stanford University researchers today released their highly anticipated 2026 AI Index Report, revealing a global landscape where artificial intelligence technology is being adopted at record-breaking

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs

SafetyDGX agent

arXiv:2604.08846v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have been shown to be vulnerable to malicious queries that can elicit unsafe responses. Recent work uses prom

From Business Events to Auditable Decisions: Ontology-Governed Graph Simulation for Enterprise AI

Model ReleasesDGX agent

arXiv:2604.08603v1 Announce Type: new Abstract: Existing LLM-based agent systems share a common architectural failure: they answer from the unrestricted knowledge space without first simulating how ac

From Paper to Program: Accelerating Quantum Many-Body Algorithm Development via a Multi-Stage LLM-Assisted Workflow

Model ReleasesDGX agent

arXiv:2604.04089v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate code rapidly but remain unreliable for scientific algorithms whose correctness depends on structural

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

SafetyDGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

I am catching glimpses in my feed that there is a backlash against Mythos as 'marketing hype,' and it is a little confusing. I don't think a…

SafetyDGX agent

I am catching glimpses in my feed that there is a backlash against Mythos as 'marketing hype,' and it is a little confusing. I don't think anyone who has used the latest agentic coding tools, would th

Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism

SafetyDGX agent

arXiv:2604.09544v1 Announce Type: cross Abstract: Large language models (LLMs) undergo alignment training to avoid harmful behaviors, yet the resulting safeguards remain brittle: jailbreaks routinely

Leave My Images Alone: Preventing Multi-Modal Large Language Models from Analyzing Images via Visual Prompt Injection

SafetyDGX agent

arXiv:2604.09024v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) have emerged as powerful tools for analyzing Internet-scale image data, offering significant benefits but al

MAB-DQA: Addressing Query Aspect Importance in Document Question Answering with Multi-Armed Bandits

SafetyDGX agent

arXiv:2604.08952v1 Announce Type: new Abstract: Document Question Answering (DQA) involves generating answers from a document based on a user's query, representing a key task in document understanding

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

Model ReleasesDGX agent

arXiv:2411.10636v2 Announce Type: replace-cross Abstract: In this study, we investigate extrinsic gender bias in Bangla pretrained language models, a largely underexplored area in low-resource languag

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

SafetyDGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.09452v1 Announce Type: cross Abstract: Safety guarantees are a prerequisite to the deployment of reinforcement learning (RL) agents in safety-critical tasks. Often, deployment environments

SynDocDis: A Metadata-Driven Framework for Generating Synthetic Physician Discussions Using Large Language Models

ApplicationsDGX agent

arXiv:2604.08555v1 Announce Type: new Abstract: Physician-physician discussions of patient cases represent a rich source of clinical knowledge and reasoning that could feed AI agents to enrich and eve

The entire team put a lot of effort into this benchmark. Document parsing and OCR certainly isn't solved, but now it is a bit easier to meas…

Model ReleasesDGX agent

The entire team put a lot of effort into this benchmark. Document parsing and OCR certainly isn't solved, but now it is a bit easier to measure 🦙 We’re open sourcing the first document OCR benchmark f

Through Their Eyes: Fixation-aligned Tuning for Personalized User Emulation

SafetyDGX agent

arXiv:2604.09368v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as scalable user simulators for recommender system evaluation. Yet existing simulators per

Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning

Model ReleasesDGX agent

arXiv:2604.08780v1 Announce Type: cross Abstract: World models promise a paradigm shift in robotics, where an agent learns the underlying physics of its environment once to enable efficient planning a

Tried Ollama Cloud, just realize only Kimi model accept images

Local AiDGX agent

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

Violence is not the answer. But maybe boycotts are?

SafetyDGX agent

Violence is not the answer. But maybe boycotts are? 🚨 NOW: The FBI is RAIDING the home of a 20-year-old man who threw a molotov cocktail at the home of OpenAI CEO Sam Altman Over a DOZEN federal agent

← Previous
1…238239240241242…297
Next →