AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
4 Jun 2026

'Your AI Text is not Mine': Redefining and Evaluating AI-generated Text Detection under Realistic Assumptions

Model ReleasesDGX agent

arXiv:2606.04906v1 Announce Type: cross Abstract: Although it is generally agreed that AI-generated text poses a broad societal risk, there is no common understanding in the AI-generated text detectio

3 Jun 2026

20x Faster Training Data Reads with Alluxio and Ray Data: A Cross-Region Benchmark

Model ReleasesDGX agent

This benchmark demonstrates how integrating Alluxio with Ray Data achieves 20x faster training data read speeds for cross-region machine learning workloads on Anyscale's platform. The study shows perf

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: P…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

95% token reduction. 30x faster execution. 90%+ task completion. Today at #MSBuild, we announced a major shift to move reasoning upstream: Pinecone Nexus now integrates directly with @Microsoft OneLak

A Benchmark for Semi-supervised Multi-modal Crowd Counting

Model ReleasesDGX agent

arXiv:2606.03646v1 Announce Type: new Abstract: This paper constructs the first benchmark on semi-supervised multi-modal crowd counting. To lay the foundation for this unexplored task, we first formul

A Single-Loop Bilevel Deep Learning Method for Optimal Control of Obstacle Problems

Model ReleasesDGX agent

arXiv:2601.04120v2 Announce Type: replace-cross Abstract: Optimal control of obstacle problems arises in a wide range of applications and is computationally challenging due to its nonsmoothness, nonli

A workflow audit is no longer the best way to figure out how to use AI in your job. Despite the advice from AI labs, I'm more convinced, bec…

Model ReleasesDGX agent

A workflow audit is no longer the best way to figure out how to use AI in your job. Despite the advice from AI labs, I'm more convinced, because of AI's reasoning capabilities and long context horizon

A^2: Smaller Self-Supervised ViTs Localize Better than Larger Ones

Local AiDGX agent

arXiv:2606.03148v1 Announce Type: new Abstract: Robust visual classification often depends on localizing the main foreground objects in an image while ignoring contextual distractors. Surprisingly, we

AdaWeather: Adaptively Mixing Probabilistic Weather Forecasts with Logarithmic Regret

ResearchDGX agent

arXiv:2606.02663v1 Announce Type: cross Abstract: Recent advances in machine learning have produced probabilistic weather forecasting models comparable to state-of-the-art numerical weather predictors

Agentic Chain-of-Thought Steering for Efficient and Controllable LLM Reasoning

AgentsDGX agent

arXiv:2606.03965v1 Announce Type: cross Abstract: Large language models improve final-answer accuracy through extended chain-of-thought reasoning, but often spend tokens inefficiently and offer little

AmbientEye: A Dataset for Pupil Segmentation under Natural Ambient Infrared Illumination

Model ReleasesDGX agent

arXiv:2606.03774v1 Announce Type: new Abstract: Eye tracking is essential for smart glasses, as it provides insight into user attention for ambient intelligence applications. However, most existing ey

Any2Poster: Any-Source Poster Generation Across Modalities and Domains

Model ReleasesDGX agent

arXiv:2606.02915v1 Announce Type: new Abstract: Visual posters are a compact medium for communicating dense information, yet progress on automatic poster generation remains difficult to measure becaus

ArrowFlow: Hierarchical Machine Learning in the Space of Permutations

Model ReleasesDGX agent

arXiv:2604.04087v2 Announce Type: replace Abstract: We introduce ArrowFlow, a machine learning architecture that operates entirely in the space of permutations. Its computational units are ranking fil

As AI gets better, it reveals an empty promise

Model ReleasesDGX agent

This week we've got tandem hands-ons with Google's new Gemini AI agent - Spark - from my colleagues David Pierce and Jay Peters. Their takeaways are similar: It's so effective that it's scary. Spark k

Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics

Model ReleasesDGX agent

arXiv:2507.21638v2 Announce Type: replace Abstract: The development of reinforcement learning (RL) algorithms has been largely driven by ambitious challenge tasks and benchmarks. Games have dominated

Auditable Climate Risk Intelligence from Fragmented ESG Data: Deterministic Orchestration and Imbalance-Aware Learning for Scope 1-3 Validation

Model ReleasesDGX agent

arXiv:2606.02604v1 Announce Type: cross Abstract: ESG and climate risk data remain fragmented across heterogeneous Scope 1, Scope 2, and Scope 3 reporting environments, while conventional validation p

Auditing Engagement Incentives in the Kidfluencer Ecosystem: A Multimodal Weak Supervision Approach

Model ReleasesDGX agent

arXiv:2606.03173v1 Announce Type: cross Abstract: The rise of `kidfluencers' on YouTube has raised ethical concerns about child digital labor and exploitation. While emerging legislation attempts to r

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

Model ReleasesDGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

b9491

Local AiDGX agent

b9491 is a release of llama.cpp , a C/C++ implementation of large language model inference that enables running LLMs locally with minimal dependencies. This release likely contains bug fixes, feature

BA-T: An Iterative Transformer for Two-View Bundle Adjustment

ResearchDGX agent

arXiv:2606.03287v1 Announce Type: new Abstract: Feed-forward models for 3D reconstruction have achieved strong performance using deep cross-view attention to exchange information across images. Howeve

Beyond 'To whom it may concern': Tailoring Machine Translation to Audience and Intent

ResearchDGX agent

arXiv:2606.03259v1 Announce Type: new Abstract: Translation quality depends on purpose: the same source text demands different translations depending on audience, tone, and communicative intent. Yet M

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs

SafetyDGX agent

arXiv:2606.03647v1 Announce Type: cross Abstract: Accurately evaluating adversarial robustness is a longstanding challenge. A flawed attack design can inflate robustness estimates, making deployment r

Building Reliable Long-Form Generation via Hallucination Rejection Sampling

ResearchDGX agent

arXiv:2606.03628v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in open-ended text generation, yet they remain prone to hallucinating incorrect or unsu

Chatbots Output Meaningful (but Problematic) Language

Model ReleasesDGX agent

arXiv:2606.02973v1 Announce Type: new Abstract: Are utterances by AI chatbots meaningful? Concretely, if a user asks, say, Anthropic's agent Claude, 'What is the capital of Spain?' and Claude answers,

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

AgentsDGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

Coherent Swap Regret and Channel-Proof Learning

Model ReleasesDGX agent

arXiv:2606.02655v1 Announce Type: cross Abstract: External regret certifies stability only against replacing one's behavior by a fixed alternative. In a quantum game, this misses a natural physical mo

Core-based Hierarchies for Efficient GraphRAG

ApplicationsDGX agent

arXiv:2603.05207v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) enhances large language models by incorporating external knowledge. However, existing vector-based method

Cost-Aware Query Routing in RAG: Empirical Analysis of Retrieval Depth Tradeoffs

Model ReleasesDGX agent

arXiv:2606.02581v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) faces a fundamental three-way tension: deeper retrieval improves factual grounding but inflates token costs and e

CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA

Model ReleasesDGX agent

arXiv:2512.00360v2 Announce Type: replace Abstract: We study timestamped question answering over educational lecture videos under a single-GPU latency/memory budget. Given a natural-language query, th

Critical evaluation of PINN for FWD inverse analysis and differentiable FEM as an alternative

Model ReleasesDGX agent

arXiv:2606.03210v1 Announce Type: cross Abstract: Automatic-differentiation-based inverse analysis methods, including physics-informed neural networks (PINNs) and differentiable programming, have rece

Demo2Tutorial: From Human Experience to Multimodal Software Tutorials

Model ReleasesDGX agent

arXiv:2606.03951v1 Announce Type: new Abstract: Human experience in digital environments offers a vast, underexplored resource of authentic, untrimmed interactions that contain rich procedural knowled

Denoise First, Orthogonalize Later: Understanding Momentum in Muon via Spectral Filtering

SafetyDGX agent

arXiv:2606.03899v1 Announce Type: new Abstract: Muon has recently demonstrated strong empirical performance in large language model training, but the theoretical role of momentum in Muon remains uncle

DeskCraft: Benchmarking Desktop Agents on Professional Workflows and Human-in-the-Loop Collaboration

Model ReleasesDGX agent

arXiv:2606.03103v1 Announce Type: new Abstract: Real-world professional desktop workflows in specialized creative and engineering software unfold over long horizons and often require human-in-the-loop

🇸🇻 El Salvador now has its own open persona dataset Today, working with NVIDIA and WideLabs, a Latin American leader in sovereign AI, we h…

Model ReleasesDGX agent

🇸🇻 El Salvador now has its own open persona dataset Today, working with NVIDIA and WideLabs, a Latin American leader in sovereign AI, we have Nemotron-Personas-El-Salvador. It’s the first open dataset

Enhanced Renewable Energy Forecasting using Context-Aware Conformal Prediction

Local AiDGX agent

arXiv:2510.15780v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) is increasingly used to support renewable energy forecasting and grid operations. As renewable penetration grows,

Enhancing Paraphrase Type Generation: The Impact of DPO and RLHF Evaluated with Human-Ranked Data

ResearchDGX agent

arXiv:2506.02018v2 Announce Type: replace Abstract: Paraphrasing re-expresses meaning to enhance applications like text simplification, machine translation, and question-answering. Specific paraphrase

EntSQL: A Benchmark for Grounding Text-to-SQL in Long-Context Enterprise Knowledge

Model ReleasesDGX agent

arXiv:2606.03363v1 Announce Type: new Abstract: Text-to-SQL enables natural language access to databases, and recent LLMs have substantially advanced its capabilities. Existing benchmarks such as Spid

EvoDrive: Pareto Evolution for Safety-Critical Autonomous Driving via Self-Improving LLM Agents

Model ReleasesDGX agent

arXiv:2606.03678v1 Announce Type: new Abstract: Generating safety-critical scenarios is essential for validating and improving autonomous driving systems, yet it inherently requires maximizing adversa

Exact equivariance, kept through training, buys zero-shot generalisation across the symmetry group

ResearchDGX agent

arXiv:2606.03003v1 Announce Type: cross Abstract: A latent world model built from an equivariant encoder E and an equivariant predictor f inherits a provable symmetry of its training loss: when the wo

Expanding Comfy With Claude Code https://x.com/i/broadcasts/1nJOLLXbrPWxR

Model ReleasesDGX agent

This X broadcast from ComfyUI likely discusses updates or new features for integrating Claude AI code capabilities with ComfyUI, a node-based UI for AI image generation and processing workflows. The s

Explainable Forecasting of Scientific Breakthroughs from Concept Network Dynamics

SafetyDGX agent

arXiv:2606.03864v1 Announce Type: cross Abstract: We introduce an explainable machine-learning approach that forecasts the structural precursors of scientific breakthroughs -- the emergence and intens

FGRPO: Federated GRPO with Adaptive Aggregation on Non-IID Data

SafetyDGX agent

arXiv:2606.03094v1 Announce Type: new Abstract: Recent advances in language models have established reinforcement learning as the primary paradigm for eliciting self-correction and long-chain reasonin

FlashbackCL: Mitigating Temporal Forgetting in Federated Learning

ResearchDGX agent

arXiv:2606.03939v1 Announce Type: cross Abstract: Federated Learning (FL) of foundation and edge models increasingly targets deployments where client data distributions drift over time, yet existing f

Flicker-DDPM: Accelerating Denoising Diffusion via 1/f Colored Noise Injection

ResearchDGX agent

arXiv:2606.03393v1 Announce Type: new Abstract: We propose a novel diffusion model, Flicker-DDPM, which incorporates flicker (1/f) noise inspired by self-organized criticality (SOC), a widely observed

Flow Learners for PDEs: Toward a Physics-to-Physics Paradigm for Scientific Computing

SafetyDGX agent

arXiv:2604.07366v2 Announce Type: replace Abstract: Partial differential equations (PDEs) govern nearly every physical process in science and engineering, but solving them at scale remains prohibitive

Forgetting is Not Erasure: Recovering Latent Knowledge via Transport Keys

SafetyDGX agent

arXiv:2606.02860v1 Announce Type: cross Abstract: Catastrophic forgetting is often framed as a representational problem: after sequential training, a model appears to lose the features that supported

Framing Migration News with LLMs: Structured CoT as a Support for Human Interpretation

Local AiDGX agent

arXiv:2606.03761v1 Announce Type: new Abstract: Frame analysis of migration news is a socially consequential task: media scholars and researchers who study how migration is narrated need tools that ar

Generating Rectifiable Measures through Neural Networks

Model ReleasesDGX agent

arXiv:2412.05109v2 Announce Type: replace Abstract: We derive universal approximation results for the class of (countably) m-rectifiable measures. Specifically, we prove that m-rectifiable measures ca

GLINT: Sparsely Gated Vision-Language Alignment for Fine-Grained Radiology Representations

SafetyDGX agent

arXiv:2606.03180v1 Announce Type: cross Abstract: Vision-language models (VLMs) for radiology have emerged as a scalable paradigm by leveraging image-report pairs naturally produced in clinical workfl

Google launches Dreambeans, an AI app that curates daily stories from Google data

Model ReleasesDGX agent

Google LLC today launched Dreambeans, an experimental app from its Google Labs division that uses artificial intelligence to assemble a finite set of personalized daily stories drawn from a user’s own

GPU-Parallel Multi-Task Reinforcement Learning with Demonstration Guided Policy Optimization

Model ReleasesDGX agent

arXiv:2606.03335v1 Announce Type: new Abstract: Large scale GPU-parallel reinforcement learning has changed what can be trained in robot simulation, yet most systems still optimize one specialist poli

Had Claude Code build a snake game where the snake becomes aware it is in the game and then... stuff happens. Some impressive creative decis…

Model ReleasesDGX agent

Had Claude Code build a snake game where the snake becomes aware it is in the game and then... stuff happens. Some impressive creative decisions by the AI (& also some very AI ones), I just gave a fir

Hierarchical Federated Learning with Dynamic Clustering and Adaptive Regularization for Robust Infrastructure Inspection

ApplicationsDGX agent

arXiv:2606.03084v1 Announce Type: new Abstract: The deployment of data-driven computer vision models for structural health monitoring (SHM) is heavily constrained by the data silo dilemma due to strin

how does the brain build and track an internal state of the world from (possibly incomplete and noisy) visual observations? i believe visual…

Model ReleasesDGX agent

how does the brain build and track an internal state of the world from (possibly incomplete and noisy) visual observations? i believe visual state tracking will be the grand challenge for vision in th

How Many Trees in a Random Forest? A Revisited Approach with Plateau Search and Optuna Integration

Model ReleasesDGX agent

arXiv:2606.03549v1 Announce Type: new Abstract: Hyperparameter optimization (HPO) for Random Forest faces a specific difficulty in tuning the number of trees: the predictive score typically improves m

https://docs.ollama.com/integrations/hermes#configure-desktop-app-with-ollama-cloud-models

Local AiDGX agent

This documentation page provides instructions for configuring the Ollama desktop application to work with Hermes models available through Ollama Cloud. It covers the setup process for integrating clou

Hybrid Adaptive Kalman Filtering for Data-Efficient Joint Tracking and Classification

ApplicationsDGX agent

arXiv:2606.02767v1 Announce Type: cross Abstract: Kalman filtering performance is highly sensitive to model mismatch and noise covariance tuning. Learning-based approaches address these limitations bu

I like the racing and Stardew Valley portions, the conclusion is very Claude, though.

Model ReleasesDGX agent

This appears to be Ethan Mollick's personal reflection on an AI-generated or AI-assisted project that includes racing and Stardew Valley game elements, with commentary that the conclusion exhibits cha

I present to you... The Spaghetti Benchmark

Model ReleasesDGX agent

'Will Smith eating spaghetti' became a shorthand for early-stage AI video generation limitations , with the original 2023 clip generated with ModelScope showing distorted faces, morphed hands, and unn

If this prompt feels well written to you, it's because Suzanne is a writer in her little spare time! You can read her short story, Mall of A…

Model ReleasesDGX agent

If this prompt feels well written to you, it's because Suzanne is a writer in her little spare time! You can read her short story, Mall of America here: https://suzannewang.com/mall-of-america It's on

In early May, the best superforecasters predicted that, by the end of the year, the longest METR 80% task horizons would reach 3-4 hours. In…

Model ReleasesDGX agent

In early May, the best superforecasters predicted that, by the end of the year, the longest METR 80% task horizons would reach 3-4 hours. In late May, Claude Mythos achieved that number. We also asked

← Previous
1…650651652653654…1042
Next →