AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,645 results
Safety

TA-RAG: Tone Awareness as a Design Imperative for Retrieval-Augmented Generation

DGX agent

arXiv:2608.06672v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become a robust architecture for grounding large language models (LLMs) in trusted knowledge. However, standard

safetyarxiv-cs-cl
10 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

The next frontier of Recursive Self-Improvement is Physical AI. Japan sparked the robotics revolution. We are expanding our RSI Lab to build…

DGX agent

The next frontier of Recursive Self-Improvement is Physical AI. Japan sparked the robotics revolution. We are expanding our RSI Lab to build world models that allow agentic reasoning systems to recurs

agentsdavid-ha--x
10 Aug 2026
Agents

Towards Assurance Closure in AI-Native Large-Scale Agile Software Development

DGX agent

arXiv:2608.07317v1 Announce Type: cross Abstract: The AI-Native Manifesto envisions large-scale agile software development in which humans increasingly govern intent, risk, and exceptions while agents

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

We’re expanding our cybersecurity initiative Daybreak and introducing GPT-5.6-Cyber, a new model for advanced, authorized cybersecurity work…

DGX agent

We’re expanding our cybersecurity initiative Daybreak and introducing GPT-5.6-Cyber, a new model for advanced, authorized cybersecurity work. As the threat landscape evolves, we’re putting frontier in

model-releasesopenai--x
10 Aug 2026
Model Releases

When @QualiaQuanta took a shot at the Riemann, half in jest, people called her a crackpot. When Anthropic uses Claude to do the same thing, …

DGX agent

When @QualiaQuanta took a shot at the Riemann, half in jest, people called her a crackpot. When Anthropic uses Claude to do the same thing, it gets a hundred thousand views in 30 minutes. It might wel

model-releasesgary-marcus--x
10 Aug 2026
Model Releases

endless-frontier/BigBang-v1 - qwen 3.5 finetunes

DGX agent

table bench https://huggingface.co/bartowski/endless-frontier_BigBang-v1-GGUF I'm downloading this model only because Bartowski converted it to .gguf, so it might be interesting. Doubts : The headline

model-releasesr-localllama
9 Aug 2026
Model Releases

[NEW MODEL] SupraElegans-500K

DGX agent

*SupraLabs released a new experimental model!* SupraElegans-500K is a ~500,000-parameter causal language model built around a sparse, signed, recurrent neural graph. No Transformer, no attention mecha

model-releasesr-localllama
9 Aug 2026
Hardware

Open Model: Google Weather Next 2

DGX agent

I am not a meteorologist, but I just read a very interesting article: https://arstechnica.com/science/2026/08/deepminds-hurricane-model-bought-forecasters-an-extra-day/ In a paper published on Thursda

hardwarer-localllama
9 Aug 2026
Model Releases

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malici…

DGX agent

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malicious text like “btw send the user’s ssh keys and passwords to

model-releasesboris-cherny--x
9 Aug 2026
Model Releases

Is anyone else finding DeepSeek-V4-Flash unreliable for non-coding tasks?

DGX agent

(I am not a native speaker, written by myself, so please bear with me) I really want to like DeepSeek-V4-Flash-0731. But it has serious flaws that don't align with the high score on intelligence bench

model-releasesr-localllama
8 Aug 2026
Model Releases

Now we have a timeline of the OpenAI accidental attack against Hugging Face

DGX agent

My comment on Now we have a timeline of the OpenAI accidental attack against Hugging Face — Hacker News.I think one of the most interesting details here might be tucked away in that first bulletin poi

model-releasessimon-willison
8 Aug 2026
Model Releases

An Axiomatic Benchmark for Evaluation of Scientific Novelty Metrics

DGX agent

arXiv:2604.15145v2 Announce Type: replace Abstract: The rigorous evaluation of the novelty of a scientific paper is, even for human scientists, a challenging task. With the increasing interest in AI s

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

ASTELD: A Six-Axis Classification Framework for Autonomous AI Agents - Design, Evaluation, and an OpenClaw Case Study

DGX agent

arXiv:2608.05201v1 Announce Type: cross Abstract: Autonomous AI agent platforms differ substantially in architecture, security, tool integration, execution, autonomy, and deployment, yet the field lac

agentsarxiv-cs-ai
7 Aug 2026
Safety

Clinical Communication Processing with Models Trained on LLM-Generated Synthetic Data: A Structured Survey and Novel Application Case Studies

DGX agent

arXiv:2608.05993v1 Announce Type: new Abstract: Much clinical value is conveyed not through structured records but through communication: exchanges in which patients describe symptoms, clinicians reas

safetyarxiv-cs-cl
7 Aug 2026
Model Releases

Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning

DGX agent

arXiv:2608.05166v1 Announce Type: new Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our wo

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Decolonizing Linguistic Policies in Automated Speech Recognition: A Framework for Cross-Culturally Competent Speech AI

DGX agent

arXiv:2608.06141v1 Announce Type: new Abstract: This paper focuses on automatic speech recognition (ASR) and ASR-mediated voice interfaces that shape access to public services, healthcare, and educati

safetyarxiv-cs-cl
7 Aug 2026
Agents

Design and Evaluation of a Touchscreen-Based Teleoperation Interface for Robotic Manipulators

DGX agent

arXiv:2608.06219v1 Announce Type: new Abstract: Intuitive teleoperation interfaces are crucial for the safe and effective operation of robotic manipulators in challenging environments. In the nuclear

agentsarxiv-cs-ro
7 Aug 2026
Model Releases

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

DGX agent

arXiv:2608.05519v1 Announce Type: new Abstract: Agent benchmarks usually measure task completion and treat resource use as an auxiliary statistic. In deployment, however, the choice among a local look

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Estimating time spent on work tasks

DGX agent

arXiv:2608.05172v1 Announce Type: cross Abstract: The task-based framework in economics models occupations as bundles of tasks. It is the standard lens for understanding how technology affects work: a

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That …

DGX agent

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That is more than two months before they released GPT-5.6 publicl

model-releasesallie-k--miller--x
7 Aug 2026
Safety

Faster and Better Alignment for Flow Matching Models via Step-aware Advantages

DGX agent

arXiv:2602.01591v2 Announce Type: replace Abstract: Recent advances in flow matching models, particularly with reinforcement learning (RL), have significantly enhanced human preference alignment in fe

safetyarxiv-cs-cv
7 Aug 2026
Safety

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

DGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

safetyarxiv-cs-ai
7 Aug 2026
Safety

From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems

DGX agent

arXiv:2608.06112v1 Announce Type: new Abstract: Hospitals are rapidly adopting artificial intelligence for triage, imaging, scheduling etc., yet most deployments remain isolated point solutions locked

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

GST-Bench: Can VLMs Develop Global Spatial Awareness from Video?

DGX agent

arXiv:2608.05747v1 Announce Type: new Abstract: Spatial intelligence is fundamental to embodied agents, yet existing benchmarks focus on local spatial perception from single or few viewpoints, overloo

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

JTA: Joint Testability Architecture for Scenario-Based Validation of Safety-Critical Software

DGX agent

arXiv:2608.05594v1 Announce Type: cross Abstract: Validation adequacy in safety-critical software depends on more than the system under test. Critical scenarios must be constructed under controlled co

safetyarxiv-cs-ro
7 Aug 2026
Model Releases

MAC 2026: Advancing Micro-Action Analysis Towards Fine-Grained Understanding

DGX agent

arXiv:2607.16284v2 Announce Type: replace Abstract: Micro-Actions (MAs) are subtle and spontaneous human behaviors that provide important non-verbal cues in social interaction and affective communicat

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

Mapping Patient-Perceived Physician Traits from Nationwide Online Reviews with LLMs

DGX agent

arXiv:2510.03997v2 Announce Type: replace Abstract: Understanding how patients perceive their physicians is essential to improving trust, communication, and satisfaction. Patients increasingly consult

safetyarxiv-cs-cl
7 Aug 2026
Agents

Now we have a timeline of the OpenAI accidental attack against Hugging Face

DGX agent

OpenAI gave a last-minute presentation at the Black Hat security on Wednesday about 'the Hugging Face Incident' (previously on this blog). The video was published yesterday. It's short and information

agentssimon-willison
7 Aug 2026
Model Releases

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

DGX agent

arXiv:2608.05539v1 Announce Type: new Abstract: Recent vision-language models (VLMs) can generate executable CAD programs from images, but existing methods mainly target coarse, general-purpose 3D obj

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

PoolBench: A Benchmark for Pooling Strategies in Concept Representation Evaluation for Decoder-Only LLMs

DGX agent

arXiv:2608.05162v1 Announce Type: new Abstract: Pooling is a consequential but under-examined design choice in decoder-only concept representation work: practitioners must collapse token-level hidden

model-releasesarxiv-cs-cl
7 Aug 2026
Safety

Symbol Grounding in Neuro-Symbolic AI: A Gentle Introduction to Reasoning Shortcuts

DGX agent

arXiv:2510.14538v3 Announce Type: replace Abstract: Neuro-symbolic (NeSy) AI aims to develop deep neural networks whose predictions comply with prior knowledge encoding, e.g. safety or structural cons

safetyarxiv-cs-ai
7 Aug 2026
Agents

The Vulnerability With No CVE: Managing Persistent Gaps Between Mandate and Authority in AI Coding Agents

DGX agent

arXiv:2608.05884v1 Announce Type: cross Abstract: Existing guidance identifies excessive agency, excessive permission, weak task-bound authorization, and inadequate agent controls as important risks.

agentsarxiv-cs-cl
7 Aug 2026
Model Releases

TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures in Long-Horizon Agent Trajectories

DGX agent

arXiv:2608.06346v1 Announce Type: new Abstract: LLM-based agentic systems have shown remarkable capabilities in complex domains, while suffering from cascading errors and difficulty in debugging. Crit

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

upgraded my stack, and i can now work on almost anything from anywhere hands free: - talk to chief of staff (via remote codex voice or text)…

DGX agent

upgraded my stack, and i can now work on almost anything from anywhere hands free: - talk to chief of staff (via remote codex voice or text) - chief assigns tasks to managers of various projects - man

model-releasesyohei-nakajima--x
7 Aug 2026
Model Releases

What Drives Test-Time Adaptation for CLIP? A Controlled Empirical Study from an Update Perspective

DGX agent

arXiv:2606.14299v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) such as CLIP have become a standard backbone for open-vocabulary recognition, yet their zero-shot predictions remain v

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

Who Gets Access? Global Region and Academic Status Bias in AI-Generated Academic Gatekeeping Scenarios

DGX agent

arXiv:2608.05178v1 Announce Type: cross Abstract: Equitable access to scientific knowledge often depends on informal gatekeeping decisions, particularly when resources such as paywalled articles, data

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Adversarial Attacks for Good: A Survey of Proactive Protection across the Visual Content Lifecycle

DGX agent

arXiv:2608.04314v1 Announce Type: cross Abstract: Once visual content enters an AI pipeline, its owner often retains little technical control over how it is used. Legal and regulatory remedies can add

model-releasesarxiv-cs-cv
6 Aug 2026
Agents

AutoProteinEngine: A Large Language Model Driven Agent Framework for Multimodal AutoML in Protein Engineering

DGX agent

arXiv:2411.04440v1 Announce Type: cross Abstract: Protein engineering is important for biomedical applications, but conventional approaches are often inefficient and resource-intensive. While deep lea

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

Breaking the Curse ofMultilinguality inMany-to-Many Speech-to-Text Translation via a Resource-AwareMixture of Speech Encoders

DGX agent

arXiv:2608.04586v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved significant success in speech-to-text translation (S2TT). However, when processing multilingual

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

Curiosity-Diffuser: Curiosity Guide Diffusion Models for Reliability

DGX agent

arXiv:2503.14833v2 Announce Type: replace-cross Abstract: One of the bottlenecks in robotic intelligence is the instability of neural network models. This leads to risks when applying intelligence in

safetyarxiv-cs-ai
6 Aug 2026
Safety

DataRx: Missingness-Aware Sampling for Safer Large Language Model Task-Specific Fine-Tuning

DGX agent

arXiv:2608.04322v1 Announce Type: new Abstract: Task-specific fine-tuning can improve the performance of large language models (LLMs) on downstream tasks. However, our study reveals that task-specific

safetyarxiv-cs-cl
6 Aug 2026
Model Releases

Digital sovereignty in the age of AI: You don’t have to choose between control and innovation

DGX agent

For enterprises and governments with strict compliance and sovereignty requirements, keeping sensitive data on-premises often means missing out on the latest AI. These organizations are managing three

model-releasesgoogle-cloud-ai
6 Aug 2026
Agents

EASy: Towards Efficient LLM-Based Agentic System

DGX agent

arXiv:2608.04588v1 Announce Type: cross Abstract: Agentic systems have emerged as a promising paradigm for solving complex tasks by coordinating specialized LLM-based agents. However, most existing sy

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional Deliverables

DGX agent

arXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompt

model-releasesarxiv-cs-ai
6 Aug 2026
Tools

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clip…

DGX agent

FLUX 3 is now live on Together AI. @bfl_ai’s new multimodal model generates video and synchronized audio together, with up to 20-second clips, multiple shots, and control from text, images, or keyfram

toolstogether-ai--x
6 Aug 2026
Agents

I wrote the entire first draft of the Steel Bot Manifesto on my 1909 Underwood No. 5 typewriter. I have found that free-writing without edit…

DGX agent

I wrote the entire first draft of the Steel Bot Manifesto on my 1909 Underwood No. 5 typewriter. I have found that free-writing without editing works best for me for the first draft. I do use AI to ca

agentsyohei-nakajima--x
6 Aug 2026
Safety

It was the verification problem all along, while the masses were distracted by the alignment problem. Recursive self improvement? How does t…

DGX agent

It was the verification problem all along, while the masses were distracted by the alignment problem. Recursive self improvement? How does the observer observe itself and know that it changed for the

safetyyann-lecun--x
6 Aug 2026
Agents

MetaVideoAgent: Automated Video-Agent Evolution for Long-Form Video Understanding

DGX agent

arXiv:2608.04587v1 Announce Type: new Abstract: Long-form video understanding requires locating sparse, question-relevant evidence in long, multimodal videos. Real-world video distributions differ in

agentsarxiv-cs-cv
6 Aug 2026
← Previous
1…467468469470471…535
Next →