AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,284 results
Model Releases

What’s new with Google Data Cloud

DGX agent

April 13 - April 17 We announced we are reintroducing Data Studio to play a significant role in the AI era, expanding from data visualizations and reports to host BigQuery conversational agents and da

model-releasesgoogle-cloud-ai
16 Apr 2026
Model Releases

When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2503.23137v2 Announce Type: replace-cross Abstract: Understanding humor-particularly when it involves complex, contradictory narratives that require comparative reasoning-remains a significant c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Wordle 1,761 3/6 ⬛⬛⬛⬛🟨 ⬛⬛🟨🟨⬛ 🟩🟩🟩🟩🟩

DGX agent

Anthropic's official X (Twitter) account shared a Wordle puzzle result for game number 1,761, solved in 3 out of 6 attempts. The post shows the characteristic colored tile pattern indicating the progr

model-releasesanthropic--x
16 Apr 2026
Model Releases

Wordle 1,762 4/6 ⬛🟨⬛⬛⬛ ⬛🟨⬛⬛⬛ 🟨⬛🟨🟨⬛ 🟩🟩🟩🟩🟩

DGX agent

This post documents a Wordle game result where the player solved puzzle #1,762 in four attempts, using the standard color-coded feedback system (gray for incorrect letters, yellow for correct letters

model-releasesanthropic--x
16 Apr 2026
Model Releases

Working Notes on Late Interaction Dynamics: Analyzing Targeted Behaviors of Late Interaction Models

DGX agent

arXiv:2603.26259v2 Announce Type: replace-cross Abstract: While Late Interaction models exhibit strong retrieval performance, many of their underlying dynamics remain understudied, potentially hiding

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

WorkRB: A Community-Driven Evaluation Framework for AI in the Work Domain

DGX agent

arXiv:2604.13055v1 Announce Type: new Abstract: Today's evolving labor markets rely increasingly on recommender systems for hiring, talent management, and workforce analytics, with natural language pr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

You can also set your autocompact threshold yourself and effectively lower your context window if you'd prefer. For example, 400k context is…

DGX agent

You can also set your autocompact threshold yourself and effectively lower your context window if you'd prefer. For example, 400k context is a good compromise: CLAUDE_CODE_AUTO_COMPACT_WINDOW=400000 c

model-releasesthariq--x
16 Apr 2026
Model Releases

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) …

DGX agent

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) consistently picking gold patches, we were sure Claude Opus

model-releasesai21-labs--x
15 Apr 2026
Model Releases

2/5 Turns out the model wasn't remembering the solution, but it was identifying 'gold-like' aesthetics like minimality & clarity. Total form…

DGX agent

AI21 Labs shared findings indicating that their model does not simply memorize solutions but instead identifies and recognizes aesthetic qualities associated with high-quality outputs, such as minimal

model-releasesai21-labs--x
15 Apr 2026
Model Releases

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents

DGX agent

arXiv:2512.20798v4 Announce Type: replace Abstract: As autonomous AI agents are deployed in high-stakes environments, ensuring their safety has become a paramount concern. Existing safety benchmarks p

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

A Foot Resistive Force Model for Legged Locomotion on Muddy Terrains

DGX agent

arXiv:2604.12006v1 Announce Type: new Abstract: Legged robots face significant challenges in moving and navigating on deformable and highly yielding terrain such as mud. We present a resistive force m

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

A General Model for Deepfake Speech Detection: Diverse Bonafide Resources or Diverse AI-Based Generators

DGX agent

arXiv:2603.27557v2 Announce Type: replace-cross Abstract: In this paper, we analyze two main factors of Bonafide Resource (BR) or AI-based Generator (AG) which affect the performance and the generalit

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

A Large-Scale Comparative Analysis of Imputation Methods for Single-Cell RNA Sequencing Data

DGX agent

arXiv:2603.24626v2 Announce Type: replace-cross Abstract: Background: Single-cell RNA sequencing (scRNA-seq) enables gene expression profiling at cellular resolution but is inherently affected by spar

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

A Layer-wise Analysis of Supervised Fine-Tuning

DGX agent

arXiv:2604.11838v1 Announce Type: cross Abstract: While critical for alignment, Supervised Fine-Tuning (SFT) incurs the risk of catastrophic forgetting, yet the layer-wise emergence of instruction-fol

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

A Sanity Check on Composed Image Retrieval

DGX agent

arXiv:2604.12904v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image, and a relative caption that specifies the

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

DGX agent

arXiv:2602.11236v2 Announce Type: replace-cross Abstract: Building general-purpose embodied agents across diverse hardware remains a central challenge in robotics, often framed as the ''one-brain, man

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Adaptive Data Dropout: Towards Self-Regulated Learning in Deep Neural Networks

DGX agent

arXiv:2604.12945v1 Announce Type: cross Abstract: Deep neural networks are typically trained by uniformly sampling large datasets across epochs, despite evidence that not all samples contribute equall

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Adobe takes Creative Cloud into Claude Code-esque territory

DGX agent

Adobe has introduced the **Firefly AI Assistant**, a new agentic tool that brings autonomous, multi-step workflow capabilities to Creative Cloud — drawing comparisons to AI coding agents like Claude C

model-releasesars-technica
15 Apr 2026
Model Releases

AffectAgent: Collaborative Multi-Agent Reasoning for Retrieval-Augmented Multimodal Emotion Recognition

DGX agent

arXiv:2604.12735v1 Announce Type: new Abstract: LLM-based multimodal emotion recognition relies on static parametric memory and often hallucinates when interpreting nuanced affective states. In this p

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

AISafetyBenchExplorer: A Metric-Aware Catalogue of AI Safety Benchmarks Reveals Fragmented Measurement and Weak Benchmark Governance

DGX agent

arXiv:2604.12875v1 Announce Type: new Abstract: The rapid expansion of large language model (LLM) safety evaluation has produced a substantial benchmark ecosystem, but not a correspondingly coherent m

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

AlphaEval: Evaluating Agents in Production

DGX agent

arXiv:2604.12162v1 Announce Type: new Abstract: The rapid deployment of AI agents in commercial settings has outpaced the development of evaluation methodologies that reflect production realities. Exi

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Analyzing the Effect of Noise in LLM Fine-tuning

DGX agent

arXiv:2604.12469v1 Announce Type: new Abstract: Fine-tuning is the dominant paradigm for adapting pretrained large language models (LLMs) to downstream NLP tasks. In practice, fine-tuning datasets may

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Anthropic rolls out identity verification that may require Claude users to provide a government-issued photo ID and live selfie to access 'certain capabilities' (Jose Antonio Lanz/Decrypt)

DGX agent

Jose Antonio Lanz / Decrypt: Anthropic rolls out identity verification that may require Claude users to provide a government-issued photo ID and live selfie to access “certain capabilities” — Anthropi

model-releasestechmeme
15 Apr 2026
Model Releases

Anthropic’s Claude Code gets automated ‘routines’ and a desktop makeover

DGX agent

Anthropic PBC is making it easier to automate tasks using Claude Code without relying on autonomous artificial intelligence agents with the launch of a new service called “routines.” The routines allo

model-releasessiliconangle
15 Apr 2026
Model Releases

AnyPoC: Universal Proof-of-Concept Test Generation for Scalable LLM-Based Bug Detection

DGX agent

arXiv:2604.11950v1 Announce Type: cross Abstract: While recent LLM-based agents can identify many candidate bugs in source code, their reports remain static hypotheses that require manual validation,

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ARC-AGI-3 has the lowest human bar of any AI benchmark out there. Almost all benchmarks require specialized knowledge that make them inacces…

DGX agent

ARC-AGI-3 has the lowest human bar of any AI benchmark out there. Almost all benchmarks require specialized knowledge that make them inaccessible to 99%+ of humans (like, say SWE-Bench). ARC-AGI-3 is

model-releasesfrancois-chollet--x
15 Apr 2026
Model Releases

Are Video Reasoning Models Ready to Go Outside?

DGX agent

arXiv:2603.10652v2 Announce Type: replace-cross Abstract: In real-world deployment, vision-language models often encounter disturbances such as weather, occlusion, and camera motion. Under such condit

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ARGOS: Who, Where, and When in Agentic Multi-Camera Person Search

DGX agent

arXiv:2604.12762v1 Announce Type: cross Abstract: We introduce ARGOS, the first benchmark and framework that reformulates multi-camera person search as an interactive reasoning problem requiring an ag

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ASTRA: Let Arbitrary Subjects Transform in Video Editing

DGX agent

arXiv:2510.01186v2 Announce Type: replace Abstract: While existing video editing methods excel with single subjects, they struggle in dense, multi-subject scenes, frequently suffering from attention d

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Banger paper from NVIDIA. Agentic reasoning needs models that are not just capable, but efficient at long-context inference. The agent model…

DGX agent

Banger paper from NVIDIA. Agentic reasoning needs models that are not just capable, but efficient at long-context inference. The agent model layer is moving toward open, long-context, high-throughput

model-releasesdair-ai--x
15 Apr 2026
Model Releases

Been waiting a month for Anthropic to answer a simple usage question about Claude Code subscriptions Have I been ghosted

DGX agent

Been waiting a month for Anthropic to answer a simple usage question about Claude Code subscriptions Have I been ghosted Can I get some questions answered by someone at Anthropic? 1. Can you use an OA

model-releasesjeremy-howard--x
15 Apr 2026
Model Releases

Benchmarking Deflection and Hallucination in Large Vision-Language Models

DGX agent

arXiv:2604.12033v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) increasingly rely on retrieval to answer knowledge-intensive multimodal questions. Existing benchmarks overlook c

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Benchmarking Foundation Models with Retrieval-Augmented Generation in Olympic-Level Physics Problem Solving

DGX agent

arXiv:2510.00919v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) with foundation models has achieved strong performance across diverse tasks, but their capacity for exper

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks

DGX agent

arXiv:2604.12379v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rely on explicit reasoning to solve coding tasks, yet evaluating the quality of this reasoning remains chall

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Perception Errors: Semantic Fixation in Large Vision-Language Models

DGX agent

arXiv:2604.12119v1 Announce Type: new Abstract: Large vision-language models (VLMs) often rely on familiar semantic priors, but existing evaluations do not cleanly separate perception failures from ru

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Beyond Relevance: On the Relationship Between Retrieval and RAG Information Coverage

DGX agent

arXiv:2603.08819v3 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems combine document retrieval with a generative model to address complex information seeking tasks l

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Scores: Diagnostic LLM Evaluation via Fine-Grained Abilities

DGX agent

arXiv:2604.12191v1 Announce Type: new Abstract: Current evaluations of large language models aggregate performance across diverse tasks into single scores. This obscures fine-grained ability variation

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Beyond Single-Dimension Novelty: How Combinations of Theory, Method, and Results-based Novelty Shape Scientific Impact

DGX agent

arXiv:2604.12471v1 Announce Type: cross Abstract: Scientific novelty drives advances at the research frontier, yet it is also associated with heightened uncertainty and potential resistance from incum

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

BID-LoRA: A Parameter-Efficient Framework for Continual Learning and Unlearning

DGX agent

arXiv:2604.12686v1 Announce Type: cross Abstract: Recent advances in deep learning underscore the need for systems that can not only acquire new knowledge through Continual Learning (CL) but also remo

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Bilevel Late Acceptance Hill Climbing for the Electric Capacitated Vehicle Routing Problem

DGX agent

arXiv:2604.13013v1 Announce Type: new Abstract: This paper tackles the Electric Capacitated Vehicle Routing Problem (E-CVRP) through a bilevel optimization framework that handles routing and charging

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Bipedal-Walking-Dynamics Model on Granular Terrains

DGX agent

arXiv:2604.11981v1 Announce Type: new Abstract: Bipeds have demonstrated high agility and mobility in unstructured environments such as sand. The yielding of such granular media brings significant sin

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

btw the famous slack chart is slack propaganda and everyone who cites it is legally obligated to also link to @sophiebits

DGX agent

btw the famous slack chart is slack propaganda and everyone who cites it is legally obligated to also link to @sophiebits Every time I see a tweet saying “I can vibe code this in a weekend” - I think

model-releasesswyx--x
15 Apr 2026
Model Releases

Budget blown on closed APIs is a solvable problem. Simply replace 20m of closed-source tokens with 1m on Minimax M2.7. Frontier performanc…

DGX agent

Budget blown on closed APIs is a solvable problem. Simply replace 20m of closed-source tokens with 1m on Minimax M2.7. Frontier performance. 20x lower cost. No rate limits. If you're rethinking your A

model-releasesfireworks-ai--x
15 Apr 2026
Model Releases

Built GPT-2, Llama 3, and DeepSeek from scratch in PyTorch - open source code + book [p]

DGX agent

A Reddit post on r/MachineLearning sharing an open-source project and accompanying book by Sebastian Raschka that walks through implementing GPT-2, Llama 3, and DeepSeek from scratch using PyTorch, wi

model-releasesr-machinelearning
15 Apr 2026
Model Releases

ByteDance launches its Seedance 2.0 video model to enterprise clients in 100+ countries, excluding the US amid legal disputes, after a February launch in China (Juro Osawa/The Information)

DGX agent

Juro Osawa / The Information: ByteDance launches its Seedance 2.0 video model to enterprise clients in 100+ countries, excluding the US amid legal disputes, after a February launch in China — ByteDanc

model-releasestechmeme
15 Apr 2026
Model Releases

Can AI Tools Transform Low-Demand Math Tasks? An Evaluation of Task Modification Capabilities

DGX agent

arXiv:2604.12743v1 Announce Type: new Abstract: While recent research has explored AI tools' ability to classify the quality of mathematical tasks (arXiv:2603.03512), little is known about their capac

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch

DGX agent

arXiv:2603.24647v4 Announce Type: replace Abstract: The autoresearch repository enables an LLM agent to optimize hyperparameters by editing training code directly. We use it as a testbed to compare cl

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

Capsule Security launches with $7M to secure AI agents at runtime

DGX agent

Israeli agentic artificial intelligence security startup Capsule Security Ltd. today launched with 7 million in new funding to expand go-to-market efforts and accelerate product development across its

model-releasessiliconangle
15 Apr 2026
← Previous
1…429430431432433…465
Next →