AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlog
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
Safety

Reinforcement Learning from Rich Feedback with Distributional DAgger

DGX agent

arXiv:2606.05152v1 Announce Type: cross Abstract: Reasoning models have advanced rapidly, but the dominant reinforcement learning from verifiable rewards (RLVR) recipe remains surprisingly narrow: sam

safetyarxiv-cs-ai
4 Jun 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Rethinking Incompleteness: Formalizing Protocol Divergence and Train-Once Learning for Robust IMVC

DGX agent

arXiv:2606.04857v1 Announce Type: new Abstract: Standard IMVC evaluation retrains separate models for different missing-data configurations. We show that this paradigm obscures a fundamental vulnerabi

researcharxiv-cs-lg
4 Jun 2026
Model Releases

Robust Multi-view Clustering against Imperfect Information

DGX agent

arXiv:2606.04343v1 Announce Type: new Abstract: Real-world multi-view data always suffer from imperfect information problem, where the view-specific observations are absent (i.e., Incomplete Views, IV

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Safety by narrow control has shown to fail many times. Need more transparency on the absolute frontier, and openness close behind.

DGX agent

Safety by narrow control has shown to fail many times. Need more transparency on the absolute frontier, and openness close behind. I found another API that offers claude-oceanus-v1-p the pricing and t

model-releasesclem-delangue--x
4 Jun 2026
Research

SANE Schema-aware Natural-language Evaluation of Biological Data

DGX agent

arXiv:2606.04500v1 Announce Type: new Abstract: High-throughput microscopy generates large, structured datasets capturing cellular responses to pharmacological perturbations, but accessing these datas

researcharxiv-cs-cl
4 Jun 2026
Safety

Scaling Self-Evolving Agents via Parametric Memory

DGX agent

arXiv:2606.04536v1 Announce Type: new Abstract: Existing memory-augmented LLM agents store past experience exclusively in prompt space, as textual summaries or retrieved passages, while keeping model

safetyarxiv-cs-ai
4 Jun 2026
Research

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots

DGX agent

arXiv:2606.04503v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has greatly advanced large reasoning models (LRMs), but it requires timely training on a huge fu

researcharxiv-cs-ai
4 Jun 2026
Model Releases

Streaming Communication in Multi-Agent Reasoning

DGX agent

arXiv:2606.05158v1 Announce Type: cross Abstract: Multi-agent reasoning systems adopt a 'generate-then-transfer' paradigm that forces end-to-end latency to scale linearly with pipeline depth. We intro

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Symbolic Regression for Shared Expressions: Introducing Partial Parameter Sharing

DGX agent

arXiv:2601.04051v3 Announce Type: replace Abstract: Symbolic regression aims to find symbolic expressions that describe datasets. Due to its inherent interpretability, symbolic regression (SR) is a po

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits so…

DGX agent

The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits society as a whole—these are exactly the principles that need

model-releasesyoshua-bengio--x
4 Jun 2026
Model Releases

The capabilities of Claude Code and Codex have expanded a lot in recent months, they added many ways to approach work (subagents, skills, go…

DGX agent

The capabilities of Claude Code and Codex have expanded a lot in recent months, they added many ways to approach work (subagents, skills, goal, workflows, plugins, etc). Given the AI labs can use thei

model-releasesethan-mollick--x
4 Jun 2026
Model Releases

The new memory system will keep track of important details automatically. If you prefer the legacy saved memories experience, you can switch…

DGX agent

The new memory system will keep track of important details automatically. If you prefer the legacy saved memories experience, you can switch back in settings. The new memory system is rolling out to P

model-releasesopenai--x
4 Jun 2026
Research

Translation Heads: Disentangling meaning from language in LLM-based machine translation

DGX agent

arXiv:2602.04613v2 Announce Type: replace Abstract: Mechanistic Interpretability (MI) seeks to explain how neural networks implement their capabilities, but the scale of Large Language Models (LLMs) h

researcharxiv-cs-cl
4 Jun 2026
Applications

UltraEP: Unleash MoE Training and Inference on Rack-Scale Nodes with Near-Optimal Load Balancing

DGX agent

arXiv:2606.04101v1 Announce Type: cross Abstract: Large-scale expert parallelism (EP) is becoming pivotal for training and serving frontier MoE models, but it also amplifies device-level expert load i

applicationsarxiv-cs-lg
4 Jun 2026
Model Releases

We're building in Canada. 🇨🇦

DGX agent

We're building in Canada. 🇨🇦 For decades, Canada invested to build the research foundations that made modern AI possible. Now we have to build, train, and scale what comes next here at home. Canada's

model-releasescohere--x
4 Jun 2026
Model Releases

We're presenting ParseBench at CVPR 2026 today. 🦙 Come learn why document understanding is an AGI-complete problem (an agent can't act on a…

DGX agent

We're presenting ParseBench at CVPR 2026 today. 🦙 Come learn why document understanding is an AGI-complete problem (an agent can't act on a doc it can't correctly read, and reading a real enterprise t

model-releasesjerry-liu--x
4 Jun 2026
Model Releases

We’ve been researching new ways for ChatGPT memory to carry context across conversations and keep it useful over time. Today, that work is r…

DGX agent

We’ve been researching new ways for ChatGPT memory to carry context across conversations and keep it useful over time. Today, that work is rolling out as a more capable memory system in ChatGPT. https

model-releasesopenai--x
4 Jun 2026
Model Releases

What Are We Actually Benchmarking in Robot Manipulation?

DGX agent

arXiv:2606.04233v1 Announce Type: new Abstract: A robotics benchmark score measures success under one fixed evaluation setup, yet is routinely treated as evidence of general manipulation capability. W

model-releasesarxiv-cs-ro
4 Jun 2026
Research

When Clients Stop Following: A Cognitive Conceptualization Diagram-driven Framework for Strategic Counseling

DGX agent

arXiv:2606.04389v1 Announce Type: new Abstract: Large Language Models (LLMs) show promise in psychological counseling, yet existing benchmarks rely heavily on highly cooperative simulated clients. We

researcharxiv-cs-cl
4 Jun 2026
Research

When Detectors Forget Forensics: Blocking Semantic Shortcuts for Generalizable AI-Generated Image Detection

DGX agent

arXiv:2603.09242v2 Announce Type: replace Abstract: The growing realism of generative models has blurred the boundary between real and synthetic content, posing significant challenges to reliable AI-g

researcharxiv-cs-cv
4 Jun 2026
Model Releases

When Do Fewer Coordinates Suffice in DP-SGD?

DGX agent

arXiv:2606.04375v1 Announce Type: new Abstract: Differentially private stochastic gradient descent (DP-SGD) injects noise into every updated coordinate, making the injected noise energy scale with the

model-releasesarxiv-cs-lg
4 Jun 2026
Local Ai

Why Muon Outperforms Adam: A Curvature Perspective

DGX agent

arXiv:2606.04662v1 Announce Type: cross Abstract: Muon improves training efficiency over Adam in large language-model training by about two times, but the local geometric source of this advantage rema

local-aiarxiv-cs-ai
4 Jun 2026
Model Releases

With the new memory system, you can review and steer what ChatGPT remembers through a memory summary, with more visibility and control over …

DGX agent

OpenAI introduced a new memory system for ChatGPT that allows users to review and control what the AI remembers across conversations through a memory summary feature. This update provides enhanced tra

model-releasesopenai--x
4 Jun 2026
Model Releases

Wordle 1,810 4/6 ⬛🟨⬛⬛⬛ ⬛⬛🟨⬛🟨 🟨🟩⬛🟨🟨 🟩🟩🟩🟩🟩

DGX agent

This entry documents a Wordle game result where the player solved puzzle #1,810 in 4 attempts, using the color-coded feedback system (black for incorrect letters, yellow for correct letters in wrong p

model-releasesanthropic--x
4 Jun 2026
Model Releases

'Your AI Text is not Mine': Redefining and Evaluating AI-generated Text Detection under Realistic Assumptions

DGX agent

arXiv:2606.04906v1 Announce Type: cross Abstract: Although it is generally agreed that AI-generated text poses a broad societal risk, there is no common understanding in the AI-generated text detectio

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

20x Faster Training Data Reads with Alluxio and Ray Data: A Cross-Region Benchmark

DGX agent

This benchmark demonstrates how integrating Alluxio with Ray Data achieves 20x faster training data read speeds for cross-region machine learning workloads on Anyscale's platform. The study shows perf

model-releasesanyscale-ray
3 Jun 2026
Model Releases

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: P…

DGX agent

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: Pinecone Nexus now integrates directly with @Microsoft OneLak

model-releasespinecone--x
3 Jun 2026
Model Releases

A Benchmark for Semi-supervised Multi-modal Crowd Counting

DGX agent

arXiv:2606.03646v1 Announce Type: new Abstract: This paper constructs the first benchmark on semi-supervised multi-modal crowd counting. To lay the foundation for this unexplored task, we first formul

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

A Single-Loop Bilevel Deep Learning Method for Optimal Control of Obstacle Problems

DGX agent

arXiv:2601.04120v2 Announce Type: replace-cross Abstract: Optimal control of obstacle problems arises in a wide range of applications and is computationally challenging due to its nonsmoothness, nonli

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

A workflow audit is no longer the best way to figure out how to use AI in your job. Despite the advice from AI labs, I'm more convinced, bec…

DGX agent

A workflow audit is no longer the best way to figure out how to use AI in your job. Despite the advice from AI labs, I'm more convinced, because of AI's reasoning capabilities and long context horizon

model-releasesallie-k--miller--x
3 Jun 2026
Local Ai

A^2: Smaller Self-Supervised ViTs Localize Better than Larger Ones

DGX agent

arXiv:2606.03148v1 Announce Type: new Abstract: Robust visual classification often depends on localizing the main foreground objects in an image while ignoring contextual distractors. Surprisingly, we

local-aiarxiv-cs-cv
3 Jun 2026
Research

AdaWeather: Adaptively Mixing Probabilistic Weather Forecasts with Logarithmic Regret

DGX agent

arXiv:2606.02663v1 Announce Type: cross Abstract: Recent advances in machine learning have produced probabilistic weather forecasting models comparable to state-of-the-art numerical weather predictors

researcharxiv-cs-ai
3 Jun 2026
Agents

Agentic Chain-of-Thought Steering for Efficient and Controllable LLM Reasoning

DGX agent

arXiv:2606.03965v1 Announce Type: cross Abstract: Large language models improve final-answer accuracy through extended chain-of-thought reasoning, but often spend tokens inefficiently and offer little

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

AmbientEye: A Dataset for Pupil Segmentation under Natural Ambient Infrared Illumination

DGX agent

arXiv:2606.03774v1 Announce Type: new Abstract: Eye tracking is essential for smart glasses, as it provides insight into user attention for ambient intelligence applications. However, most existing ey

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Any2Poster: Any-Source Poster Generation Across Modalities and Domains

DGX agent

arXiv:2606.02915v1 Announce Type: new Abstract: Visual posters are a compact medium for communicating dense information, yet progress on automatic poster generation remains difficult to measure becaus

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

ArrowFlow: Hierarchical Machine Learning in the Space of Permutations

DGX agent

arXiv:2604.04087v2 Announce Type: replace Abstract: We introduce ArrowFlow, a machine learning architecture that operates entirely in the space of permutations. Its computational units are ranking fil

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

As AI gets better, it reveals an empty promise

DGX agent

This week we've got tandem hands-ons with Google's new Gemini AI agent - Spark - from my colleagues David Pierce and Jay Peters. Their takeaways are similar: It's so effective that it's scary. Spark k

model-releasesthe-verge-ai
3 Jun 2026
Model Releases

Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics

DGX agent

arXiv:2507.21638v2 Announce Type: replace Abstract: The development of reinforcement learning (RL) algorithms has been largely driven by ambitious challenge tasks and benchmarks. Games have dominated

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Auditable Climate Risk Intelligence from Fragmented ESG Data: Deterministic Orchestration and Imbalance-Aware Learning for Scope 1-3 Validation

DGX agent

arXiv:2606.02604v1 Announce Type: cross Abstract: ESG and climate risk data remain fragmented across heterogeneous Scope 1, Scope 2, and Scope 3 reporting environments, while conventional validation p

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Auditing Engagement Incentives in the Kidfluencer Ecosystem: A Multimodal Weak Supervision Approach

DGX agent

arXiv:2606.03173v1 Announce Type: cross Abstract: The rise of `kidfluencers' on YouTube has raised ethical concerns about child digital labor and exploitation. While emerging legislation attempts to r

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

DGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

b9491

DGX agent

b9491 is a release of llama.cpp , a C/C++ implementation of large language model inference that enables running LLMs locally with minimal dependencies. This release likely contains bug fixes, feature

local-aillama-cpp-releases
3 Jun 2026
Research

BA-T: An Iterative Transformer for Two-View Bundle Adjustment

DGX agent

arXiv:2606.03287v1 Announce Type: new Abstract: Feed-forward models for 3D reconstruction have achieved strong performance using deep cross-view attention to exchange information across images. Howeve

researcharxiv-cs-cv
3 Jun 2026
Research

Beyond 'To whom it may concern': Tailoring Machine Translation to Audience and Intent

DGX agent

arXiv:2606.03259v1 Announce Type: new Abstract: Translation quality depends on purpose: the same source text demands different translations depending on audience, tone, and communicative intent. Yet M

researcharxiv-cs-cl
3 Jun 2026
Safety

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs

DGX agent

arXiv:2606.03647v1 Announce Type: cross Abstract: Accurately evaluating adversarial robustness is a longstanding challenge. A flawed attack design can inflate robustness estimates, making deployment r

safetyarxiv-cs-ai
3 Jun 2026
Research

Building Reliable Long-Form Generation via Hallucination Rejection Sampling

DGX agent

arXiv:2606.03628v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in open-ended text generation, yet they remain prone to hallucinating incorrect or unsu

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Chatbots Output Meaningful (but Problematic) Language

DGX agent

arXiv:2606.02973v1 Announce Type: new Abstract: Are utterances by AI chatbots meaningful? Concretely, if a user asks, say, Anthropic's agent Claude, 'What is the capital of Spain?' and Claude answers,

model-releasesarxiv-cs-cl
3 Jun 2026
Agents

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

DGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

agentsarxiv-cs-ai
3 Jun 2026
← Previous
1…822823824825826…1316
Next →