The changes presented in technical reports and papers on Kimi K3, Cosmos 3, Omni, and others are not direct wins or losses on the same metric. Rather, they show that models are beginning to process more types of information within more unified architectures, extending from text and images to video, audio, actions, 3D geometry, or latent representations. Publicly available materials support this expansion in architectural scope; however, they use different tasks, metrics, and evaluation systems, which are not yet sufficient to demonstrate an overall performance leap confirmed by a unified benchmark.
Read articleIn August 2026, four automation vendors—Nintex, Resolve, UiPath, and Orkes—released or updated their orchestration platforms for AI agents within two weeks. Sources: Nintex Source: Resolve Source: UiPath Source: Orkes This is no coincidence. Together, they respond to an ongoing reality: enterprise AI is shifting from 'assistive Copilots' to 'autonomous workflows,' yet most companies are stuck in the middle, unable to let Copilots act independently or to adapt automation scripts to complex, evolving business conditions.
Read articleIn August 2026, a study published in the journal Science sparked debate: junior developers are the heaviest users of generative AI coding tools, yet they reap little benefit; instead, it is senior developers who leverage these tools to boost productivity and innovation. Daniotti and colleagues trained a machine learning classifier to identify AI-generated code among software developers in six countries—the United States, China, France, Germany, India, and Russia—and drew their conclusions from that.
Read articleAI chip 'cost breakthroughs' often conflate two separate accounting ledgers: the investment required to train a model, and the ongoing expense of keeping that trained model in service. The former involves long-cycle cluster computing, data preparation, engineering, and infrastructure; the latter is more directly affected by response latency, energy efficiency, and per-inference service cost. Improvements in inference chip efficiency metrics cannot be taken to imply that the total cost of training frontier models has dropped by the same magnitude.
Read articleA study published in Science reports that large language models outperformed a baseline of several hundred physicians in five highly challenging clinical case reasoning experiments. The study also compared second opinions from human experts and AI on randomly selected patients in the emergency department of a large tertiary academic medical center. The authors concluded that the models have surpassed most clinical reasoning benchmarks and, therefore, require urgent prospective trials.
Read articleIn February 2025, the first batch of provisions under the EU's Artificial Intelligence Act came into force, marking the implementation phase of the world's first comprehensive AI regulation. The Act adopts a risk-based classification for AI systems, imposing transparency, safety, and information security obligations on general-purpose AI (GPAI) models, and explicitly prohibits eight categories of AI uses deemed to pose "unacceptable risks." Against this framework, how are tech companies translating regulatory requirements into concrete actions? From Microsoft and OpenAI to Hugging Face, enterprises with different roles have charted distinct paths of response.
Read article