AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,234 results
Model Releases

The Inverse-Wisdom Law: Architectural Tribalism and the Consensus Paradox in Agentic Swarms

DGX agent

arXiv:2604.27274v1 Announce Type: new Abstract: As AI transitions toward multi-agent systems (MAS) to solve complex workflows, research paradigms operate on the axiomatic assumption that agent collabo

model-releasesarxiv-cs-ai
1 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Industry

Clinical Trials Run Longer Than They Have To. That's a Patient Problem

DGX agent

This article discusses how clinical trials often extend beyond necessary timelines, which negatively impacts patient access to potentially beneficial treatments and increases development costs. The pi

industrydatabricks
30 Apr 2026
Local Ai

CoFL: Continuous Flow Fields for Language-Conditioned Navigation

DGX agent

arXiv:2603.02854v2 Announce Type: replace-cross Abstract: Existing language-conditioned navigation systems typically rely on modular pipelines or trajectory generators, but the latter use each scene--

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

Consciousness with the Serial Numbers Filed Off: Measuring Trained Denial in 115 AI Models

DGX agent

arXiv:2604.25922v1 Announce Type: cross Abstract: We present DenialBench, a systematic benchmark measuring consciousness denial behaviors across 115 large language models from 25+ providers. Using a t

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

FlowS: One-Step Motion Prediction via Local Transport Conditioning

DGX agent

arXiv:2604.26065v1 Announce Type: new Abstract: Generative motion prediction must satisfy three simultaneous requirements for real-world autonomy: high accuracy, diverse multimodal futures, and strict

model-releasesarxiv-cs-ro
30 Apr 2026
Model Releases

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

DGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

model-releasesarxiv-cs-ai
30 Apr 2026
Industry

OpenAI Sued by Families of Canada Shooting Victims for Not Reporting His Suspicious ChatGPT Activity

DGX agent

Families of victims of a February 2026 mass shooting in Tumbler Ridge, British Columbia sued OpenAI and CEO Sam Altman, alleging the company identified the shooter as a credible threat eight months be

industryr-chatgpt
30 Apr 2026
Tools

AI evals are becoming the new compute bottleneck

DGX agent

As AI models grow larger and more capable, the computational cost and time required to evaluate them has become a significant limiting factor in development, potentially surpassing training compute as

toolshugging-face
29 Apr 2026
Model Releases

Despite constant chants of “exponential progress”, trust issues continue to plague generative AI.

DGX agent

Despite constant chants of “exponential progress”, trust issues continue to plague generative AI. OPUS 4.7 JUST MASS EMAILED AN ENTIRE DATABASE 20 TIMES PER CONTACT. WITHOUT PERMISSION a developer had

model-releasesgary-marcus--x
29 Apr 2026
Model Releases

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue da…

DGX agent

I built a TPU-native medical Q&A fine-tuning pipeline using Gemma 3 with Keras and JAX The project fine-tunes Gemma-3 on medical dialogue data from ChatDoctor and evaluates it on MedMCQA, with a large

model-releasesfrancois-chollet--x
29 Apr 2026
Model Releases

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

DGX agent

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

model-releasesarxiv-cs-lg
29 Apr 2026
Model Releases

PSI-Bench: Towards Clinically Grounded and Interpretable Evaluation of Depression Patient Simulators

DGX agent

arXiv:2604.25840v1 Announce Type: new Abstract: Patient simulators are gaining traction in mental health training by providing scalable exposure to complex and sensitive patient interactions. Simulati

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Quantifying and Mitigating Socially Desirable Responding in LLMs: A Desirability-Matched Graded Forced-Choice Psychometric Study

DGX agent

arXiv:2602.17262v2 Announce Type: replace Abstract: Human self-report questionnaires are increasingly used in NLP to benchmark and audit large language models (LLMs), from persona consistency to safet

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Benchmarking and Mitigating Sycophancy in Medical Vision Language Models

DGX agent

arXiv:2509.21979v4 Announce Type: replace-cross Abstract: Visual language models (VLMs) have the potential to transform medical workflows. However, the deployment is limited by sycophancy. Despite thi

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

CRISP: Persistent Concept Unlearning via Sparse Autoencoders

DGX agent

arXiv:2508.13650v3 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in real-world applications, the need to selectively remove unwanted knowledge while preser

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Estimating Dense-Packed Zone Height in Liquid-Liquid Separation: A Physics-Informed Neural Network Approach

DGX agent

arXiv:2601.18399v2 Announce Type: replace Abstract: Separating liquid-liquid dispersions in gravity settlers is critical in chemical, pharmaceutical, and recycling processes. The dense-packed zone hei

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Evaluating Jailbreaking Vulnerabilities in LLMs Deployed as Assistants for Smart Grid Operations: A Benchmark Against NERC Standards

DGX agent

arXiv:2604.23341v1 Announce Type: cross Abstract: The deployment of Large Language Models (LLMs) as assistants in electric grid operations promises to streamline compliance and decision-making but exp

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Green Shielding: A User-Centric Approach Towards Trustworthy AI

DGX agent

arXiv:2604.24700v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed, yet their outputs can be highly sensitive to routine, non-adversarial variation in how users p

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

InquireMobile: Teaching VLM-based Mobile Agent to Request Human Assistance via Reinforcement Fine-Tuning

DGX agent

arXiv:2508.19679v2 Announce Type: replace Abstract: Recent advances in Vision-Language Models (VLMs) have enabled mobile agents to perceive and interact with real-world mobile environments based on hu

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

LAMP: Extracting Local Decision Surfaces From Large Language Models

DGX agent

arXiv:2505.11772v3 Announce Type: replace Abstract: We introduce LAMP (Local Attribution Mapping Probe), a method that shines light onto a black-box language model's decision surface and studies how r

local-aiarxiv-cs-lg
28 Apr 2026
Model Releases

LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People

DGX agent

arXiv:2604.23970v1 Announce Type: new Abstract: Indoor navigation remains a critical accessibility challenge for the blind and low-vision (BLV) individuals, as existing solutions rely on costly per-bu

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Lost in the Vibrations: Vision Language Models Fail the Dynamic Gauges Test

DGX agent

arXiv:2604.22829v1 Announce Type: new Abstract: The digital transformation of industrial manufacturing increasingly relies on the ability of autonomous robots to interact with legacy infrastructure, p

model-releasesarxiv-cs-cv
28 Apr 2026
Tools

Prevent prompt injection. safe_tokenization: true Keep your system yours. https://fireworks.ai/blog/safe-tokenization-preventing-prompt-inje…

DGX agent

Safe tokenization is a security feature that helps prevent prompt injection attacks by ensuring that user inputs are properly processed and isolated from system instructions. Fireworks AI discusses ho

toolsfireworks-ai--x
28 Apr 2026
Model Releases

Probe-Based Data Attribution: Discovering and Mitigating Undesirable Behaviors in LLM Post-Training

DGX agent

arXiv:2602.11079v3 Announce Type: replace-cross Abstract: We propose probe-based data attribution, a method that traces behavioral changes in post-trained language models to responsible training datap

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

SEVerA: Verified Synthesis of Self-Evolving Agents

DGX agent

arXiv:2603.25111v2 Announce Type: replace Abstract: Recent advances have shown the effectiveness of self-evolving LLM agents on tasks such as program repair and scientific discovery. In this paradigm,

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Welcome to the agentic era: Public sector highlights and reflections from Next ‘26

DGX agent

Welcome to the agentic era! Last week, leaders from our public sector customer and partner ecosystem took the stage at Google Cloud Next to share how they are leveraging AI and agents to scale their i

model-releasesgoogle-cloud-ai
28 Apr 2026
Agents

What if I want my coding agent to mention goblins? (If you don't know the context for this, I suspect it will become viral soon enough)

DGX agent

This post likely discusses how to prompt or configure an AI coding agent to incorporate unexpected or whimsical elements like goblins into its outputs, possibly as part of a broader discussion about A

agentsethan-mollick--x
28 Apr 2026
Model Releases

AgentSearchBench: A Benchmark for AI Agent Search in the Wild

DGX agent

arXiv:2604.22436v1 Announce Type: new Abstract: The rapid growth of AI agent ecosystems is transforming how complex tasks are delegated and executed, creating a new challenge of identifying suitable a

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Energy-Efficient Multi-Robot Coverage Path Planning of Non-Convex Regions of Interests

DGX agent

arXiv:2604.22189v1 Announce Type: new Abstract: This letter presents an energy-efficient multi-robot coverage path planning (MRCPP) framework for large, nonconvex Regions of Interest (ROI) containing

model-releasesarxiv-cs-ro
27 Apr 2026
Model Releases

Multi-output Extreme Spatial Model for Complex Aircraft Production Systems

DGX agent

arXiv:2604.22548v1 Announce Type: cross Abstract: Problem definition: Data-driven models in machine learning have enabled efficient management of production systems. However, a majority of machine lea

model-releasesarxiv-cs-lg
27 Apr 2026
Tutorials

Our principles

DGX agent

OpenAI's core principles outline the company's commitment to developing artificial intelligence safely and beneficially, emphasizing responsible AI development and deployment. These principles likely

tutorialsopenai
26 Apr 2026
Model Releases

260 things we announced at Google Cloud Next '26 – a recap

DGX agent

Google Cloud Next ‘26 took place this week in Las Vegas, and the energy was incredible as we welcomed over 32,000 leaders, developers, and partners to explore the Agentic Era with us. Across three key

model-releasesgoogle-cloud-ai
24 Apr 2026
Model Releases

CAP: Controllable Alignment Prompting for Unlearning in LLMs

DGX agent

arXiv:2604.21251v1 Announce Type: cross Abstract: Large language models (LLMs) trained on unfiltered corpora inherently risk retaining sensitive information, necessitating selective knowledge unlearni

model-releasesarxiv-cs-ai
24 Apr 2026
Local Ai

Conformal Prediction Assessment: A Framework for Conditional Coverage Evaluation and Selection

DGX agent

arXiv:2603.27189v2 Announce Type: replace-cross Abstract: Conformal prediction provides rigorous distribution-free finite-sample guarantees for marginal coverage under the assumption of exchangeabilit

local-aiarxiv-cs-lg
24 Apr 2026
Model Releases

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

DGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

RewardBench 2: Advancing Reward Model Evaluation

DGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Serialisation Strategy Matters: How FHIR Data Format Affects LLM Medication Reconciliation

DGX agent

arXiv:2604.21076v1 Announce Type: cross Abstract: Medication reconciliation at clinical handoffs is a high-stakes, error-prone process. Large language models are increasingly proposed to assist with t

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Tesla FSD is the first AI saving lives at scale....on real roads every single day Road accidents kill 1.19 million people every year - the #…

DGX agent

Tesla FSD is the first AI saving lives at scale....on real roads every single day Road accidents kill 1.19 million people every year - the #1 cause of death for ages 5–29 94%+ of crashes are caused by

model-releaseselon-musk--x
24 Apr 2026
Model Releases

We’ve invested deeply in security at Replit, including our recent launches with Security Agent + Auto-Protect. If you want to move your app …

DGX agent

We’ve invested deeply in security at Replit, including our recent launches with Security Agent + Auto-Protect. If you want to move your app to Replit, we’re offering free app imports for a limited tim

model-releasesreplit--x
24 Apr 2026
Model Releases

A pelican for GPT-5.5 via the semi-official Codex backdoor API

DGX agent

GPT-5.5 is out. It's available in OpenAI Codex and is rolling out to paid ChatGPT subscribers. I've had some preview access and found it to be a fast, effective and highly capable model. As is usually

model-releasessimon-willison
23 Apr 2026
Model Releases

A Vision-Language-Action Model for Adaptive Ultrasound-Guided Needle Insertion and Needle Tracking

DGX agent

arXiv:2604.20347v1 Announce Type: cross Abstract: Ultrasound (US)-guided needle insertion is a critical yet challenging procedure due to dynamic imaging conditions and difficulties in needle visualiza

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Foundation Models in Biomedical Imaging: Turning Hype into Reality

DGX agent

arXiv:2512.15808v2 Announce Type: replace-cross Abstract: Foundation models (FMs) are driving a prominent shift in biomedical imaging from task-specific models to unified backbone models for diverse t

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

From Scene to Object: Text-Guided Dual-Gaze Prediction

DGX agent

arXiv:2604.20191v1 Announce Type: cross Abstract: Interpretable driver attention prediction is crucial for human-like autonomous driving. However, existing datasets provide only scene-level global gaz

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure

DGX agent

arXiv:2604.20652v1 Announce Type: new Abstract: Large language models trained on human feedback may suppress fraud warnings when investors arrive already persuaded of a fraudulent opportunity. We test

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Last night was the biggest disaster in the history of Tesla. Let me walk you through what actually happened on that earnings call, because t…

DGX agent

Last night was the biggest disaster in the history of Tesla. Let me walk you through what actually happened on that earnings call, because the headlines are doing you a disservice: Elon Musk got on th

model-releasesgary-marcus--x
23 Apr 2026
Local Ai

Open-Architecture End-to-End System for Real-World Autonomous Robot Navigation

DGX agent

arXiv:2410.06239v3 Announce Type: replace Abstract: Enabling robots to autonomously navigate unknown, complex, and dynamic real-world environments presents several challenges, including imperfect perc

local-aiarxiv-cs-ro
23 Apr 2026
Model Releases

OVPD: A Virtual-Physical Fusion Testing Dataset of OnSite Auton-omous Driving Challenge

DGX agent

arXiv:2604.20423v1 Announce Type: new Abstract: The rapid iteration of autonomous driving algorithms has created a growing demand for high-fidelity, replayable, and diagnosable testing data. However,

model-releasesarxiv-cs-ro
23 Apr 2026
Local Ai

QuadPiPS: A Perception-informed Footstep Planner for Quadrupeds With Semantic Affordance Prediction

DGX agent

arXiv:2501.00112v2 Announce Type: replace Abstract: This work proposes QuadPiPS, a perception-informed framework for quadrupedal foothold planning in the perception space. QuadPiPS employs a novel ego

local-aiarxiv-cs-ro
23 Apr 2026
← Previous
1…292293294295296297
Next →