AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,292 results
Model Releases

A Proactive EMR Assistant for Doctor-Patient Dialogue: Streaming ASR, Belief Stabilization, and Preliminary Controlled Evaluation

DGX agent

arXiv:2604.13059v1 Announce Type: new Abstract: Most dialogue-based electronic medical record (EMR) systems still behave as passive pipelines: transcribe speech, extract information, and generate the

model-releasesarxiv-cs-cl
16 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

a quick fix if you saw higher rate limit usage in Opus 4.7 today- hope you enjoy trying it out

DGX agent

a quick fix if you saw higher rate limit usage in Opus 4.7 today- hope you enjoy trying it out We fixed a bug where rate limits on Claude subscriptions weren't properly adjusted for long context reque

model-releasesthariq--x
16 Apr 2026
Model Releases

A Study of Failure Modes in Two-Stage Human-Object Interaction Detection

DGX agent

arXiv:2604.13448v1 Announce Type: new Abstract: Human-object interaction (HOI) detection aims to detect interactions between humans and objects in images. While recent advances have improved performan

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Accelerating the cyber defense ecosystem that protects us all

DGX agent

OpenAI discusses efforts to strengthen collective cybersecurity defenses through collaboration and technology advancement within the broader cyber defense ecosystem. The article likely covers OpenAI's

model-releasesopenai
16 Apr 2026
Model Releases

Adaptive Multi-Scale Channel-Spatial Attention Aggregation Framework for 3D Indoor Semantic Scene Completion Toward Assisting Visually Impaired

DGX agent

arXiv:2602.16385v4 Announce Type: replace Abstract: Independent indoor mobility remains a critical challenge for individuals with visual impairments, largely due to the limited capability of existing

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Aerial Vision-Language Navigation with a Unified Framework for Spatial, Temporal and Embodied Reasoning

DGX agent

arXiv:2512.08639v3 Announce Type: replace Abstract: Aerial Vision-and-Language Navigation (VLN) aims to enable unmanned aerial vehicles (UAVs) to interpret natural language instructions and navigate c

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

AeTHERON: Autoregressive Topology-aware Heterogeneous Graph Operator Network for Fluid-Structure Interaction

DGX agent

arXiv:2604.13369v1 Announce Type: cross Abstract: Surrogate modeling of body-driven fluid flows where immersed moving boundaries couple structural dynamics to chaotic, unsteady fluid phenomena remains

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, …

DGX agent

Agent evals are drifting away from production reality. Most benchmarks use clean tasks, well-specified requirements, deterministic metrics, and retrospective curation. Production work is messier, with

model-releasesdair-ai--x
16 Apr 2026
Model Releases

AI Powered Image Analysis for Phishing Detection

DGX agent

arXiv:2604.13555v1 Announce Type: new Abstract: Phishing websites now rely heavily on visual imitation-copied logos, similar layouts, and matching colours-to avoid detection by text- and URL-based sys

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Alibaba unveils Qwen3.6-35B-A3B, an open-weight MoE model with 35B total and 3B active parameters, saying it rivals larger dense models in agentic coding tasks (Qwen)

DGX agent

Qwen: Alibaba unveils Qwen3.6-35B-A3B, an open-weight MoE model with 35B total and 3B active parameters, saying it rivals larger dense models in agentic coding tasks — · 4355 words · QwenTeam丨Translat

model-releasestechmeme
16 Apr 2026
Model Releases

also available on the Claude Blog: https://claude.com/blog/using-claude-code-session-management-and-1m-context

DGX agent

Claude Code's session management capabilities and 1 million token context window are highlighted in this post, which references an official Anthropic blog entry. The feature allows developers to maint

model-releasesthariq--x
16 Apr 2026
Model Releases

also now available on the Claude Blog: https://claude.com/blog/using-claude-code-session-management-and-1m-context

DGX agent

Claude Code's session management features and 1 million token context window capabilities are now documented on the official Claude Blog. The post, shared by Thariq on X, covers how developers can lev

model-releasesthariq--x
16 Apr 2026
Model Releases

Amazon launches its first smart warehouse in Shenzhen, aiming to cut local merchant storage costs by up to 45% as competition with Shein and Temu intensifies (Iris Deng/South China Morning Post)

DGX agent

Iris Deng / South China Morning Post: Amazon launches its first smart warehouse in Shenzhen, aiming to cut local merchant storage costs by up to 45% as competition with Shein and Temu intensifies — Am

model-releasestechmeme
16 Apr 2026
Model Releases

An Empirical Investigation of Practical LLM-as-a-Judge Improvement Techniques on RewardBench 2

DGX agent

arXiv:2604.13717v1 Announce Type: new Abstract: LLM-as-a-judge, using a language model to score or rank candidate responses, is widely used as a scalable alternative to human evaluation in RLHF pipeli

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learning

DGX agent

arXiv:2211.16780v3 Announce Type: replace-cross Abstract: In online incremental learning, data continuously arrives with substantial distributional shifts, creating a significant challenge because pre

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Analog Optical Inference on Million-Record Mortgage Data

DGX agent

arXiv:2604.13251v1 Announce Type: new Abstract: Analog optical computers promise large efficiency gains for machine learning inference, yet no demonstration has moved beyond small-scale image benchmar

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Anthropic launches Claude Opus 4.7 with coding, visual reasoning improvements

DGX agent

Anthropic PBC today opened access to Claude Opus 4.7, the latest addition to its popular line of large language models. The company says that the LLM is significantly better than its predecessor at co

model-releasessiliconangle
16 Apr 2026
Model Releases

Anthropic releases a new Opus model amid Mythos Preview buzz

DGX agent

Anthropic has released its most powerful 'generally available' model to date: Claude Opus 4.7. The company called it a step up from Opus 4.6 for advanced software engineering tasks, particularly in co

model-releasesthe-verge-ai
16 Apr 2026
Model Releases

Anthropic says Opus 4.7 hits 80.6% on Document Reasoning — up from 57.1%. But 'reasoning about documents' ≠ 'parsing documents for agents.' …

DGX agent

Anthropic says Opus 4.7 hits 80.6% on Document Reasoning — up from 57.1%. But 'reasoning about documents' ≠ 'parsing documents for agents.' We ran it on ParseBench. → Charts: 13.5% → 55.8% (+42.3) — h

model-releasesjerry-liu--x
16 Apr 2026
Model Releases

As agentic AI overwhelms enterprise defenses, Oracle makes the case for security baked into the database

DGX agent

As organizations transition from experimentation with AI to full-scale production, the demand for mission-critical security and absolute data availability has become the primary benchmark for enterpri

model-releasessiliconangle
16 Apr 2026
Model Releases

ASTER: Latent Pseudo-Anomaly Generation for Unsupervised Time-Series Anomaly Detection

DGX agent

arXiv:2604.13924v1 Announce Type: cross Abstract: Time-series anomaly detection (TSAD) is critical in domains such as industrial monitoring, healthcare, and cybersecurity, but it remains challenging d

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

ASTRA: Enhancing Multi-Subject Generation with Retrieval-Augmented Pose Guidance and Disentangled Position Embedding

DGX agent

arXiv:2604.13938v1 Announce Type: new Abstract: Subject-driven image generation has shown great success in creating personalized content, but its capabilities are largely confined to single subjects i

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

AudioX: A Unified Framework for Anything-to-Audio Generation

DGX agent

arXiv:2503.10522v4 Announce Type: replace-cross Abstract: Audio and music generation based on flexible multimodal control signals is a widely applicable topic, with the following key challenges: 1) a

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Auto-FP: An Experimental Study of Automated Feature Preprocessing for Tabular Data

DGX agent

arXiv:2310.02540v2 Announce Type: replace Abstract: Classical machine learning models, such as linear models and tree-based models, are widely used in industry. These models are sensitive to data dist

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Autonomous Multi-objective Alloy Design through Simulation-guided Optimization

DGX agent

arXiv:2507.16005v2 Announce Type: replace-cross Abstract: Alloy discovery is constrained by vast compositional spaces, competing objectives, and prohibitive experimental costs. Although simulations an

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

BenGER: A Collaborative Web Platform for End-to-End Benchmarking of German Legal Tasks

DGX agent

arXiv:2604.13583v1 Announce Type: new Abstract: Evaluating large language models (LLMs) for legal reasoning requires workflows that span task design, expert annotation, model execution, and metric-bas

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Beyond Static Personas: Situational Personality Steering for Large Language Models

DGX agent

arXiv:2604.13846v1 Announce Type: new Abstract: Personalized Large Language Models (LLMs) facilitate more natural, human-like interactions in human-centric applications. However, existing personalizat

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Beyond Uniform Sampling: Synergistic Active Learning and Input Denoising for Robust Neural Operators

DGX agent

arXiv:2604.13316v1 Announce Type: new Abstract: Neural operators have emerged as fast surrogate models for physics simulations, yet they remain acutely vulnerable to adversarial perturbations, a criti

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Breaking the Generator Barrier: Disentangled Representation for Generalizable AI-Text Detection

DGX agent

arXiv:2604.13692v1 Announce Type: new Abstract: As large language models (LLMs) generate text that increasingly resembles human writing, the subtle cues that distinguish AI-generated content from huma

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

DGX agent

arXiv:2604.13472v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized c

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Built an political benchmark for LLMs. KIMI K2 can't answer about Taiwan (Obviously). GPT-5.3 refuses 100% of questions when given an opt-out. [P]

DGX agent

A researcher on r/MachineLearning built a political benchmark to evaluate how various LLMs handle sensitive geopolitical and politically contentious questions. Key findings include that Kimi K2 (Moons

model-releasesr-machinelearning
16 Apr 2026
Model Releases

Can Large Language Models Reliably Extract Physiology Index Values from Coronary Angiography Reports?

DGX agent

arXiv:2604.13077v1 Announce Type: new Abstract: Coronary angiography (CAG) reports contain clinically relevant physiological measurements, yet this information is typically in the form of unstructured

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

CANVAS: Continuity-Aware Narratives via Visual Agentic Storyboarding

DGX agent

arXiv:2604.13452v1 Announce Type: new Abstract: Long-form visual storytelling requires maintaining continuity across shots, including consistent characters, stable environments, and smooth scene trans

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Chain of Uncertain Rewards with Large Language Models for Reinforcement Learning

DGX agent

arXiv:2604.13504v1 Announce Type: cross Abstract: Designing effective reward functions is a cornerstone of reinforcement learning (RL), yet it remains a challenging and labor-intensive process due to

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon auton…

DGX agent

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon autonomy, unlocking a class of deep investigation work we couldn'

model-releasescognition-ai--x
16 Apr 2026
Model Releases

Claude Opus 4.7 is now available in Cursor. We've found it to be impressively autonomous and more creative in its reasoning. We're launching…

DGX agent

Cursor has integrated Claude Opus 4.7 into its platform, highlighting the model's impressive autonomous capabilities and enhanced creative reasoning abilities. The announcement suggests a new feature

model-releasescursor--x
16 Apr 2026
Model Releases

Claude Opus 4.7 is now available in Windsurf 2.0! Anthropic has clearly optimized Claude Opus 4.7 for sustained reasoning over long runs. Ag…

DGX agent

Claude Opus 4.7 is now available in Windsurf 2.0! Anthropic has clearly optimized Claude Opus 4.7 for sustained reasoning over long runs. Agents stay on track longer without intervention, so engineers

model-releaseswindsurf--x
16 Apr 2026
Model Releases

Claude Opus 4.7 is now the default orchestration model powering Computer. It's also available for Max subscribers on Perplexity web, iOS, an…

DGX agent

Claude Opus 4.7 has been set as the default orchestration model for Anthropic's Computer product. The model is also available to Max subscribers on Perplexity's web platform and iOS application.

model-releasesperplexity--x
16 Apr 2026
Model Releases

Claude remains irreducibly Claude. If you know, you know. (The fact that models have distinct personalities that are consistent across gener…

DGX agent

Claude remains irreducibly Claude. If you know, you know. (The fact that models have distinct personalities that are consistent across generations is technically interesting, it also makes it very eas

model-releasesethan-mollick--x
16 Apr 2026
Model Releases

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

DGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Codex for (almost) everything

DGX agent

OpenAI's Codex is a large language model trained on publicly available code from the internet that can understand and generate code in dozens of programming languages. It powers GitHub Copilot and can

model-releasesopenai
16 Apr 2026
Model Releases

Coherence in the brain unfolds across separable temporal regimes

DGX agent

arXiv:2512.20481v4 Announce Type: replace-cross Abstract: To maintain coherence in language, the brain must satisfy key competing temporal demands: the gradual accumulation of meaning across extended

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generation

DGX agent

arXiv:2604.13946v1 Announce Type: cross Abstract: Automated code generation remains a persistent challenge in software engineering, as conventional multi-agent frameworks are often constrained by stat

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Common to Whom? Regional Cultural Commonsense and LLM Bias in India

DGX agent

arXiv:2601.15550v3 Announce Type: replace Abstract: Existing cultural commonsense benchmarks treat nations as monolithic, assuming uniform practices within national boundaries. But does cultural commo

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Correct Chains, Wrong Answers: Dissociating Reasoning from Output in LLM Logic

DGX agent

arXiv:2604.13065v1 Announce Type: new Abstract: LLMs can execute every step of chain-of-thought reasoning correctly and still produce wrong final answers. We introduce the Novel Operator Test, a bench

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Correct Prediction, Wrong Steps? Consensus Reasoning Knowledge Graph for Robust Chain-of-Thought Synthesis

DGX agent

arXiv:2604.14121v1 Announce Type: new Abstract: LLM reasoning traces suffer from complex flaws -- *Step Internal Flaws* (logical errors, hallucinations, etc.) and *Step-wise Flaws* (overthinking, unde

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Counterfactual Peptide Editing for Causal TCR--pMHC Binding Inference

DGX agent

arXiv:2604.13256v1 Announce Type: new Abstract: Neural models for TCR-pMHC binding prediction are susceptible to shortcut learning: they exploit spurious correlations in training data -- such as pepti

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Covariance-adapting algorithm for semi-bandits with application to sparse rewards

DGX agent

arXiv:2604.13738v1 Announce Type: cross Abstract: We investigate stochastic combinatorial semi-bandits, where the entire joint distribution of outcomes impacts the complexity of the problem instance (

model-releasesarxiv-cs-lg
16 Apr 2026
← Previous
1…425426427428429…465
Next →