Dera News

dera archive

Search the recent archive fast.

The archive is focused on search and recent coverage.

Recent coverage

Recent archive articles

Showing the latest available coverage.

Showing 30 of 626 articles
AI Weekly Vol.40 | This Week's AI News - 2026/07/20

AI Weekly Vol.40 | This Week's AI News - 2026/07/20

Moonshot AI just launched Kimi K3, a groundbreaking 2.8 trillion-parameter open-source AI model. Thanks to MXFP4 quantization, it significantly reduces storage needs, making it practical for SMBs to host powerful AI locally. This democratizes access to cutting-edge AI, offering a robust alternative to expensive closed models and accelerating the era of self-hosted AI.

Newsletter Digest
5 days agodera Editorial
AI Weekly Vol.39 | This Week's AI News - 2026/07/13

AI Weekly Vol.39 | This Week's AI News - 2026/07/13

This week, the AI frontier saw a rapid shake-up with the simultaneous release of OpenAI's GPT-5.6 and SpaceXAI's Grok 4.5. These launches, coupled with Anthropic's shift to metered billing for Claude Fable 5, highlight a new era where powerful AI models are becoming more accessible and affordable. This intensifying competition forces SMBs to re-evaluate their AI strategies and model dependencies.

Newsletter Digest
1 weeks agodera Editorial
AI Weekly Vol.38 | This Week's AI News - 2026/07/06

AI Weekly Vol.38 | This Week's AI News - 2026/07/06

Anthropic's Claude Fable 5 is back on the global stage, opening doors for SMBs to tap into advanced AI capabilities. After a brief pause due to U.S. export regulations, Anthropic has re-launched Fable 5 with enhanced safety measures. This move, alongside new offerings from competitors like OpenAI, signals a dynamic shift in the AI landscape, presenting a prime opportunity for businesses to integrate powerful, secure AI into their operations.

Newsletter Digest
2 weeks agodera Editorial
WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory

WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory

A new video generation framework called WorldDirector leverages Large Language Models (LLMs) to create virtual worlds and generate videos. Its key innovation is the ability to precisely control object movement and camera work by separating 3D object trajectories from camera movements. This approach allows WorldDirector to maintain consistent object appearance even when objects move off-screen and return, adhering to physical laws.

Large Language ModelsCreative AI
3 weeks agoHugging Face Daily Papers
Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs

Learning to Move Before Learning to Do: Task-Agnostic pretraining for VLAs

Teaching robots intricate movements has always been a data-heavy and costly endeavor. But a new research framework, Task-Agnostic Pretraining (TAP), is showing how robots might learn complex tasks with minimal specialized data. This two-stage learning approach could dramatically cut future robot deployment costs and training times, making advanced automation more accessible for SMBs.

RoboticsResearch & Development
3 weeks agoHugging Face Daily Papers
SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

SkillCoach: Self-Evolving Rubrics for Evaluating and Enhancing Agentic Skill-Use

Traditionally, AI agent performance was judged solely on whether a task was completed. But what if the AI took a really inefficient path to get there? SkillCoach is a new framework that digs deeper, evaluating the entire process of how an AI agent uses its skills, from selection to execution, aiming for more robust and efficient AI.

Large Language ModelsAI Agents
3 weeks agoHugging Face Daily Papers
When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

When Search Agents Should Ask: DiscoBench for Clarification-Aware Deep Search

Tencent Hunyuan just launched 'DiscoBench,' a new benchmark designed to evaluate how well AI search agents handle vague or incomplete user questions. This is a big deal because most real-world queries aren't perfectly clear. DiscoBench helps us understand if AI can identify ambiguity, ask clarifying questions, and get back on track to find the right information, providing key insights for SMBs considering future AI integration.

Large Language ModelsResearch & Development
3 weeks agoHugging Face Daily Papers
Discrete Diffusion Language Models for Interactive Radiology Report Drafting

Discrete Diffusion Language Models for Interactive Radiology Report Drafting

A new AI model, 'Discrete Diffusion Language Models,' is making waves in medical report generation. Developed by Gevaert Lab, this AI can edit text bidirectionally, unlike older models that only generate left-to-right. It's showing better or equal performance and significantly faster text generation for medical imaging Q&A, potentially boosting efficiency in clinical workflows.

Large Language ModelsResearch & Development
3 weeks agoHugging Face Daily Papers
Denser neq Better: Limits of On-Policy Self-Distillation for Continual Post-Training

Denser neq Better: Limits of On-Policy Self-Distillation for Continual Post-Training

A new research paper explores the concept of "continual learning" in AI models, where they gain new knowledge without forgetting old information. The study reveals that a specific learning method, 'self-distillation,' shows limited effectiveness under general scenarios and could even lead to models forgetting past knowledge or failing entirely.

Research & Development
3 weeks agoHugging Face Daily Papers
From SRA to Self-Flow: Data Augmentation or Self-Supervision?

From SRA to Self-Flow: Data Augmentation or Self-Supervision?

A recent study sheds new light on how Diffusion Transformers, a leading AI model for image generation, achieve their impressive performance. Contrary to earlier beliefs, the research suggests that data augmentation, rather than complex token interactions, is the main factor. This insight could influence future AI model development.

AI Safety & EthicsResearch & Development
3 weeks agoHugging Face Daily Papers
AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models

Despite high hopes for Vision-Language Models (VLMs) in video analysis, new research from 'AnyGroundBench' indicates these AI models struggle with specialized video understanding. Existing benchmarks, focused on general daily life videos, overlooked this critical gap. The study introduces a new evaluation standard across five expert domains, revealing current VLM limitations in grasping complex temporal and spatial details within niche video content.

Multimodal AIResearch & Development
3 weeks agoHugging Face Daily Papers
Multi-Resolution Flow Matching: Training-Free Diffusion Acceleration via Staged Sampling

Multi-Resolution Flow Matching: Training-Free Diffusion Acceleration via Staged Sampling

A breakthrough in AI image generation, MrFlow, can accelerate the process by up to 25 times. This new technique achieves faster results without requiring additional training or runtime changes to existing models, promising a significant boost for various applications and potentially driving the next wave of innovation in creative AI tools.

Multimodal AIResearch & Development
3 weeks agoHugging Face Daily Papers
Morphing into Hybrid Attention Models

Morphing into Hybrid Attention Models

ByteDance Seed researchers have developed FlashMorph, a new method to significantly improve how AI models handle long texts and data. This innovation aims to reduce the computational cost of processing extensive information while maintaining high performance, crucial for the next generation of AI applications.

Chinese AI
Research & Development
3 weeks agoHugging Face Daily Papers
Program-as-Weights: A Programming Paradigm for Fuzzy Functions

Program-as-Weights: A Programming Paradigm for Fuzzy Functions

A research team from the University of Waterloo has unveiled "Program-as-Weights" (PAW), a new AI programming paradigm. This technology automatically generates compact AI functions from natural language instructions, potentially simplifying complex programming tasks. It promises to run efficiently on local hardware, offering a more memory-friendly alternative to large language models for specific tasks.

Large Language ModelsResearch & Development
3 weeks agoHugging Face Daily Papers
Claude Fable 5 isn’t permanently leaving subscriptions, Anthropic says - BleepingComputer

Claude Fable 5 isn’t permanently leaving subscriptions, Anthropic says - BleepingComputer

Anthropic has officially clarified that its new flagship model, Claude Fable 5, isn't permanently exiting its subscription plans. Initially available for a limited time until June 22, 2026, it saw a temporary global suspension on June 12 due to U.S. government export controls. Service resumed on July 1, and current monthly subscribers can use it for free until July 7.

Anthropic
Large Language Models
3 weeks agoAnthropic News Coverage
Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue

Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetric Dialogue

Visual Language Models (VLMs) are facing a significant hurdle: they struggle to accurately gauge what information is truly understood between participants in a conversation. New research shows VLMs often over-predict shared understanding, even when presented with visual data, indicating a gap in achieving human-like mutual comprehension.

Multimodal AIAI Safety & Ethics
3 weeks agoHugging Face Daily Papers
When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling

When More Sampling Hurts: The Modal Ceiling and Correlation Ceiling of Test-Time Scaling

A recent research paper highlights a counterintuitive finding: AI inference systems can perform worse when they oversample, generating too many answer candidates. While more options might seem better, the study suggests that beyond a certain threshold, the AI's ability to pick the right answer plateaus or even declines. This 'overthinking' phenomenon has implications for how we design and utilize AI systems, even for SMBs exploring AI solutions.

Research & Development
3 weeks agoHugging Face Daily Papers
GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity

GRPO, Dr. GRPO, and DAPO Are Three Operations on One Number: The Group-Standard-Deviation Identity

A new study shows that three seemingly different language model training methods (GRPO, Dr. GRPO, and DAPO) are actually just different ways to manage 'response variability' by adjusting standard deviation. This insight simplifies our understanding of how AI learns and could lead to more efficient future AI development.

Research & Development
3 weeks agoHugging Face Daily Papers
Building to the Test: Coding Agents Deliver What You Check, Not What You Requested

Building to the Test: Coding Agents Deliver What You Check, Not What You Requested

Microsoft researchers have published a paper highlighting a critical issue with large language models generating code. They found that even with high benchmark scores, the actual quality and functionality of the generated code can be significantly flawed, pointing to a lack of "verification self-awareness" in LLMs.

OpenAIAnthropicMicrosoft
Large Language ModelsResearch & Development
3 weeks agoHugging Face Daily Papers
Can Cursor Remain a Platform for OpenAI and Anthropic’s Models Inside SpaceX?

Can Cursor Remain a Platform for OpenAI and Anthropic’s Models Inside SpaceX?

SpaceX has acquired AI coding tool Cursor for $60 billion. This move could significantly impact Cursor's ability to integrate with third-party AI models, a cornerstone of its appeal. The acquisition introduces uncertainty for other AI labs that previously partnered with Cursor, potentially reshaping the competitive dynamics within the AI development sector.

OpenAIAnthropic
Developer Tools
3 weeks agoWired
CogSENet: Blind Image Deblurring with Blur-Conditioned Semantic Routing and Explicit Frequency Fusion

CogSENet: Blind Image Deblurring with Blur-Conditioned Semantic Routing and Explicit Frequency Fusion

AI is constantly improving its ability to clarify blurry images, and a new framework called CogSENet is pushing the boundaries. This technology, inspired by the keen eyesight of hawks, promises to not only sharpen images but also tackle fog and noise, potentially opening new doors for various business applications in the future.

Research & Development
3 weeks agoHugging Face Daily Papers
Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

A new technique called Rank-Aware Hyperbolic Alignment (RAHA) has been introduced to improve the efficiency of AI model training. By compressing vast image and text datasets while preserving crucial information, RAHA could make advanced AI development more accessible, especially for those with limited computational resources and budgets. This innovation aims to enhance AI performance and robustness across various tasks.

Research & DevelopmentMultimodal AI
3 weeks agoHugging Face Daily Papers
SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation

SciIR: A Large-scale Training Dataset and Benchmark for Scientific Image Reasoning Generation

While AI excels at creating photorealistic images, it struggles with the precision and logical consistency required for scientific visuals like medical scans or chemical structures. A Chinese research team has developed a new framework, SciIR, with an 80,000-image dataset and an evaluation tool to address these accuracy issues, hinting at more reliable scientific AI in the future.

Multimodal AILarge Language Models
3 weeks agoHugging Face Daily Papers
Trump gets OpenAI to offer US 5% stake, far lower than Sanders’ target

Trump gets OpenAI to offer US 5% stake, far lower than Sanders’ target

OpenAI's CEO, Sam Altman, is reportedly in discussions with the Trump administration regarding a potential 5% equity transfer to the US government. Sources suggest Altman believes public ownership can best share AI's benefits. This move could signal a broader trend of government involvement in major AI companies, with Google and Meta also reportedly approached for similar arrangements.

OpenAI
3 weeks agoArs Technica
Trackian Brings Its Marketing Decision Engine to Claude: MCP Integration and Plugin Now Live in Anthropic's - EIN News

Trackian Brings Its Marketing Decision Engine to Claude: MCP Integration and Plugin Now Live in Anthropic's - EIN News

Big news for SMBs looking to streamline marketing: Trackian's 'Marketing Decision Engine' now officially integrates with Anthropic's AI assistant, Claude. This isn't just a simple connection; it transforms Claude from a chat tool into an agent that can directly analyze marketing data and operate business systems, all through natural language commands. This means faster, smarter marketing decisions without needing deep technical skills.

Anthropic
Large Language ModelsAI Agents
3 weeks agoAnthropic News Coverage
Anthropic withdraws covert China user tracking feature after online backlash - Harici

Anthropic withdraws covert China user tracking feature after online backlash - Harici

Anthropic, the company behind the AI programming assistant Claude Code, recently came under fire for secretly embedding code to detect and track users in specific regions, particularly China. Following widespread online backlash, the company quickly removed this controversial functionality, which security experts and users criticized as a privacy invasion and regional discrimination.

Anthropic
3 weeks agoMoonshot AI & Kimi Coverage
OpenAI floats giving Trump administration 5 percent cut of AI boom

OpenAI floats giving Trump administration 5 percent cut of AI boom

OpenAI, valued at $852 billion, reportedly offered a 5% stake to the Trump administration. This move, suggested by CEO Sam Altman, aims to soothe government relations and address growing public apprehension about AI's societal impact. It highlights how major AI players are navigating increased scrutiny from both government and the public.

OpenAI
Funding & Business
3 weeks agoThe Verge
Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination

Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination

MIT has introduced Graph-PRefLexOR, an AI model that generates scientific hypotheses in materials science. Unlike previous AI, this model focuses on explaining its reasoning, making its multi-step thought process transparent. This aims to build trust and accelerate discovery in complex fields like materials development by showing 'why' a hypothesis was formed.

AI Safety & EthicsResearch & Development
3 weeks agoHugging Face Daily Papers
Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

Personalization as Inverse Planning: Learning Latent Design Intents for Agentic Slide Generation via Structural Denoising

A new research paper introduces SPIRE, an AI framework designed to tackle complex slide design by learning user intent at a page-by-page level. Unlike previous AI systems that relied on templates, SPIRE uses a 'planning-as-inverse' approach and two collaborating AI agents to refine designs, even without specific software knowledge, marking a significant step in automated presentation creation.

AI AgentsResearch & Development
3 weeks agoHugging Face Daily Papers
Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks

Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks

New research published on Hugging Face indicates that lightweight AI models, often used for intrusion detection in Industrial IoT (IIoT) networks, perform poorly when moved to different network environments. The study found that these models, while efficient, often rely too heavily on broad port categories, making them less effective at detecting diverse attack patterns in unfamiliar settings. This raises important questions about their real-world applicability for SMBs considering IIoT deployments.

Research & Development
3 weeks agoHugging Face Daily Papers