AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,630 results
Model Releases

BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts

DGX agent

arXiv:2606.10061v1 Announce Type: new Abstract: Large language models (LLMs) increasingly participate in emotionally sensitive social conversations, where responses may shift from balanced support tow

model-releasesarxiv-cs-cl
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Beyond Model Size: Probing the Gaps in Visual in-Context Learning by Training a Tiny Model

DGX agent

arXiv:2606.10905v1 Announce Type: new Abstract: Visual in-Context Learning (VICL) aims at making progress towards adaptive vision models, that can -- based on a few examples -- adapt to a new task at

model-releasesarxiv-cs-cv
10 Jun 2026
Applications

BiWM: Advancing Open-Source Interactive Video World Models with Bidirectional Autoregression

DGX agent

arXiv:2606.10135v1 Announce Type: cross Abstract: Transitioning bidirectional video diffusion models into an autoregressive paradigm improves the interactivity of video world models, but existing caus

applicationsarxiv-cs-ai
10 Jun 2026
Applications

Co-GLANCE: Uncertainty-Aware Active Perception for Heterogeneous Robot Teaming

DGX agent

arXiv:2606.09919v1 Announce Type: cross Abstract: Perceptual uncertainty is a central challenge for heterogeneous robot teams operating in unstructured outdoor environments, where no single viewpoint

applicationsarxiv-cs-ai
10 Jun 2026
Tutorials

Concentration of power, capabilities and economic wealth is the biggest risk in AI. We need open science and open-source more than ever!

DGX agent

Jeremy Howard argues that the concentration of AI power, capabilities, and economic wealth among few entities represents the most significant risk in AI development, and advocates for open science and

tutorialsjeremy-howard--x
10 Jun 2026
Model Releases

Cyst-X: A Multi-Center MRI Benchmark and Federated Learning Framework for Malignancy-Risk Stratification of Pancreatic Cystic Neoplasm

DGX agent

arXiv:2507.22017v4 Announce Type: replace-cross Abstract: Pancreatic cancer is projected to be the second-deadliest cancer by 2030, making early detection critical. Intraductal papillary mucinous neop

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation

DGX agent

arXiv:2606.10142v1 Announce Type: new Abstract: Recent advances in 3D generation have led to substantial improvements in realism, controllability, and efficiency, yet the evaluation of 3D assets remai

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

DiffusionGemma

DGX agent

DiffusionGemma Last May Google briefly released an experimental Gemini Diffusion model. I tried the preview at the time and recorded it running at 857 tokens/second. It was an exciting model, but Goog

model-releasessimon-willison
10 Jun 2026
Model Releases

Don't waste SAM

DGX agent

arXiv:2606.10696v1 Announce Type: new Abstract: Meta AI has recently released the Segment Anything Model (SAM), which demonstrates exceptional zero-shot image segmentation performance across various t

model-releasesarxiv-cs-cv
10 Jun 2026
Safety

Enhancing Multilingual LLM-based ASR with Mixture of Experts and Dynamic Downsampling

DGX agent

arXiv:2606.10439v1 Announce Type: cross Abstract: The rapid progress of large language models (LLMs) has opened up a new frontier for automatic speech recognition (ASR), making their effective integra

safetyarxiv-cs-cl
10 Jun 2026
Applications

From Senses to Decisions: The Information Flow of Auditory and Visual Perception in Multimodal LLMs

DGX agent

arXiv:2606.10147v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) can listen and see, but how do audio and visual signals actually travel through the network to shape an answer?

applicationsarxiv-cs-ai
10 Jun 2026
Model Releases

GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines

DGX agent

arXiv:2606.09935v1 Announce Type: cross Abstract: AI-powered agents are increasingly embedded in continuous integration and continuous delivery/deployment (CI/CD) pipelines to autonomously review pull

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

IDP-Bench: Benchmarking ability of LLMs to protect personal information in interdependent privacy contexts

DGX agent

arXiv:2606.09908v1 Announce Type: cross Abstract: Large language models (LLMs) are becoming widely deployed as personal AI assistants with access to sensitive user data, making privacy a major challen

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

🔔 IF Anthropic is for real about safety, the should pause, for one month, and show leadership. If they won’t pause, even briefly, we have a…

DGX agent

🔔 IF Anthropic is for real about safety, the should pause, for one month, and show leadership. If they won’t pause, even briefly, we have a different kind of answer about what is real and what is mark

safetygary-marcus--x
10 Jun 2026
Model Releases

If Claude Fable stops helping you, you'll never know

DGX agent

If Claude Fable stops helping you, you'll never know Jonathon Ready highlights one of the more eyebrow-raising details from the 319 page system card for Fable 5 and Mythos 5. Here's a longer excerpt,

model-releasessimon-willison
10 Jun 2026
Model Releases

IPSM-Bench: A New Intermediate Phase Segmentation Benchmark in Microstructure Images of Zinc-Based Absorbable Biomaterials

DGX agent

arXiv:2606.11001v1 Announce Type: new Abstract: Zinc-based alloys are indispensable emerging absorbable metallic biomaterials, and their macroscopic performance is governed by microstructural characte

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

Janus: A Benchmark for Goal-Conditioned Information Distortion in LLMs

DGX agent

arXiv:2606.10852v1 Announce Type: cross Abstract: LLM deception is often evaluated through direct markers such as fabricated claims, explicit lies, or strategic concealment. However, many real-world m

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

mlr3mbo: Bayesian Optimization in R

DGX agent

arXiv:2603.29730v2 Announce Type: replace-cross Abstract: We present mlr3mbo, a modular toolbox for Bayesian optimization in R. mlr3mbo supports single- and multi-objective optimization, multi-point p

model-releasesarxiv-cs-lg
10 Jun 2026
Safety

One paper from three years ago has been influencing policymakers around the world about labour decisions regarding AI. What happens when the…

DGX agent

One paper from three years ago has been influencing policymakers around the world about labour decisions regarding AI. What happens when the gap between the evidence and new policy widens? Our report

safetycohere--x
10 Jun 2026
Model Releases

PhantomBench: Benchmarking the Non-existential Threat of Language Models

DGX agent

arXiv:2606.11105v1 Announce Type: cross Abstract: Hallucinations, where language models (LMs) generate factually ungrounded responses, pose serious risks, as users tend to blindly rely on them. This i

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

DGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Quoting Jeremy Howard

DGX agent

Easy solution to slow down recursive AI self improvement: The lab with the top-ranked model must agree THEY must not use it for working on frontier AI But everyone else should have access to it. By de

model-releasessimon-willison
10 Jun 2026
Industry

Racist comments targeting politicians tripled since Meta relaxed its rules

DGX agent

After Meta relaxed its content moderation policies in January 2025, analysis of nearly 8 million Facebook comments showed that abusive and racist comments targeting lawmakers from both parties tripled

industryars-technica
10 Jun 2026
Applications

Rod models in continuum and soft robot control: a review

DGX agent

arXiv:2407.05886v3 Announce Type: replace Abstract: Continuum and soft robots can transform automation tasks requiring compliant interaction in constrained or unstructured environments, including heal

applicationsarxiv-cs-ro
10 Jun 2026
Model Releases

T1-Bench: Benchmarking Multi-Scenario Agents in Real-World Domains

DGX agent

arXiv:2606.11070v1 Announce Type: cross Abstract: Recent advances in reasoning and tool-calling capabilities of large language models (LLMs) have enabled increasingly capable agentic systems. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Applications

Toward Proactive RF Charging Scheduling: Generative AI for Decision Support

DGX agent

arXiv:2606.10600v1 Announce Type: cross Abstract: Radio frequency wireless power transfer (RF-WPT) is an enabling technology for supporting uninterrupted communications in future Internet of Things sy

applicationsarxiv-cs-lg
10 Jun 2026
Agents

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

DGX agent

arXiv:2606.10749v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly moving from conversational interfaces to software components that plan, invoke tools, maintain memory, a

agentsarxiv-cs-ai
10 Jun 2026
Agents

Understanding and mitigating the risks of OpenClaw for non-technical users: A practical guide with Skill

DGX agent

arXiv:2606.11007v1 Announce Type: cross Abstract: OpenClaw has rapidly emerged as a transformative artificial intelligence (AI) agent framework, and its ability to autonomously execute complex, multi-

agentsarxiv-cs-ai
10 Jun 2026
Safety

Warren to SEC, lightly paraphrased: “Do your f’ing job, and don’t let retail investors get screwed”

DGX agent

Senator Elizabeth Warren criticized the SEC for insufficient enforcement and investor protection, urging the agency to strengthen oversight and prevent harm to retail investors. The post, shared by AI

safetygary-marcus--x
10 Jun 2026
Model Releases

What makes a harness a harness: necessary and sufficient conditions for an agent harness

DGX agent

arXiv:2606.10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence. It names the layer that wraps a language

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models

DGX agent

arXiv:2606.10740v1 Announce Type: new Abstract: Failures in multi-turn reasoning models are largely invisible to terminal-score evaluation. A model can lock onto an unsafe stance early in a long dialo

safetyarxiv-cs-ai
10 Jun 2026
Safety

When you hear AI 'safety' you should hear 'censorship' and 'control' instead. All of us surveilled and spied by safeguards of loving grace. …

DGX agent

When you hear AI 'safety' you should hear 'censorship' and 'control' instead. All of us surveilled and spied by safeguards of loving grace. Today it's intelligent Terms of Service control. You can't d

safetyyann-lecun--x
10 Jun 2026
Model Releases

Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields

DGX agent

arXiv:2606.11042v1 Announce Type: new Abstract: Recent years have witnessed the rapid evolution of AI agents toward handling increasingly complex, real-world tasks. However, existing benchmarks rarely

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

A Framework for Evaluating and Benchmarking Concept Drift Detection Methods

DGX agent

arXiv:2606.07789v1 Announce Type: new Abstract: Data stream mining is fundamentally challenged by concept drift, where distributional changes can degrade model performance. Despite the proliferation o

model-releasesarxiv-cs-lg
9 Jun 2026
Safety

Agent Economics: An Entropy-Controlled Pluralistic Alignment Framework for Preventing Artificial Hivemind in Autonomous Agents

DGX agent

arXiv:2606.09039v1 Announce Type: new Abstract: This study proposes the Behavioral Protocol Framework (BPF), an entropy-controlled pluralistic alignment framework designed to address two critical chal

safetyarxiv-cs-ai
9 Jun 2026
Local Ai

An Alternative Trajectory for Generative AI

DGX agent

arXiv:2603.14147v2 Announce Type: replace Abstract: The generative artificial intelligence (AI) ecosystem is undergoing rapid transformations that threaten its sustainability. As models transition fro

local-aiarxiv-cs-ai
9 Jun 2026
Industry

Apple says its AI is still private, even when it's running on Google's servers

DGX agent

Apple is expanding its Private Cloud Compute (PCC) beyond its data centers, partnering with Google and NVIDIA to run Apple Intelligence workloads on Google Cloud. Apple says it worked with Google and

industryars-technica
9 Jun 2026
Safety

Autonomous FPV Flight with Translational Optical Flow and Uncertainty Mask

DGX agent

arXiv:2606.09088v1 Announce Type: new Abstract: Autonomous FPV quadrotor flight in complex environments using a monocular RGB camera as the sole exteroceptive sensor remains a fundamental challenge. R

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

Benchmark Datasets for Lead-Lag Forecasting on Social Platforms

DGX agent

arXiv:2511.03877v2 Announce Type: replace Abstract: Social and collaborative platforms emit multivariate time-series traces in which early interactions -- such as views, likes, or downloads -- are fol

model-releasesarxiv-cs-lg
9 Jun 2026
Hardware

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly …

DGX agent

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly degrade its IQ so that the average engineer won't notice. We

hardwareclem-delangue--x
9 Jun 2026
Tools

Can Voice Agents Handle Bilingual Customers? Benchmarking Frontier ASR on Code-Switched Speech

DGX agent

This paper benchmarks frontier automatic speech recognition (ASR) systems on their ability to handle code-switched speech, where bilingual speakers mix languages within a single conversation. The rese

toolshugging-face
9 Jun 2026
Safety

Causal Transfer in Medical Image Analysis

DGX agent

arXiv:2603.24388v2 Announce Type: replace Abstract: Medical imaging models frequently fail when deployed across hospitals, scanners, populations, or imaging protocols due to domain shift, limiting the

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage Sites in China

DGX agent

arXiv:2606.08959v1 Announce Type: new Abstract: We introduce ChinaHeritaQA, a multimodal benchmark dataset for evaluating the cultural reasoning abilities of vision-language models (VLMs) on UNESCO Wo

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

@cohere Nice! Great to see another open source model released. 🙌

DGX agent

Cohere announced the release of another open source model, receiving positive reception from the community. The post was shared on X (formerly Twitter) and highlights Cohere's continued contribution t

model-releasescohere--x
9 Jun 2026
Safety

Comparative evaluation of training strategies using partially labelled datasets for segmentation of white matter hyperintensities and stroke lesions in FLAIR MRI

DGX agent

arXiv:2601.20503v2 Announce Type: replace-cross Abstract: White matter hyperintensities (WMH) and ischaemic stroke lesions (ISL) are key imaging biomarkers of cerebral small vessel disease (SVD) detec

safetyarxiv-cs-ai
9 Jun 2026
Safety

Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing

DGX agent

arXiv:2606.07636v1 Announce Type: new Abstract: Editing a long-form video from heterogeneous footage requires more than selecting clips: an agent must preserve narrative intent across material prepara

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

Cross-View Urban Traffic Dataset: Drone-Supervised Ground Truth for Monocular Bird's-Eye View Localization

DGX agent

arXiv:2606.07708v1 Announce Type: cross Abstract: We introduce a dataset and benchmark for cross-view urban traffic perception built from synchronized ego-centric bicycle videos and aerial drone video

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Deep reinforcement learning for process design: Review and perspective

DGX agent

arXiv:2308.07822v2 Announce Type: replace Abstract: The transformation towards renewable energy and feedstock supply in the chemical industry requires new conceptual process design approaches. Recentl

agentsarxiv-cs-lg
9 Jun 2026
← Previous
1…488489490491492…534
Next →