
Claude Opus 5がARC-AGI-3で満点、NVIDIAが示した「ハーネス」設計の重要性
NVIDIAは自社基盤AVOで動かしたClaude Opus 5がARC-AGI-3の公開セットで満点を達成したと発表。単体計測(別条件)は約30%だったとしており、モデルではなくハーネス設計が実力を左右する構造転換と、NVIDIAのGPU需要拡大という狙いを解説する。
Apache Sparkの創設者らによって設立されたデータ分析プラットフォーム企業。データウェアハウスとデータレイクの利点を組み合わせた「レイクハウス」アーキテクチャを提唱し、企業のデータ活用とAI導入を支援する。

NVIDIAは自社基盤AVOで動かしたClaude Opus 5がARC-AGI-3の公開セットで満点を達成したと発表。単体計測(別条件)は約30%だったとしており、モデルではなくハーネス設計が実力を左右する構造転換と、NVIDIAのGPU需要拡大という狙いを解説する。

Databricksの社内ベンチマークで、トークン単価が安いSonnet 5のタスク単価がOpus 4.8を上回る逆転現象が判明した。完了率とハーネス選定がコストを左右する構造をUberの予算超過事例とあわせて読み解く。

NVIDIAがAI企業のKumo AIを買収した。同社のグラフニューラルネットワーク技術により、企業は膨大なデータから顧客行動や需要を容易に予測可能となる。NVIDIAはこの技術を統合し、AIインフラから業務予測までを担う垂直統合戦略を加速させる狙いだ。

OpenAIはCodexに6種の役割別プラグイン、Sites、アノテーションを追加し、開発者向けのコード支援から部門横断の業務成果物生成基盤へ広げる。非開発者利用の拡大が、その転換を支えている。

Claude Opus 4.8が掲げる最大の進化は「正直さ」だ。自分が書いたコードの欠陥を見逃す確率は前世代の約4分の1に下がった。一方でAnthropicは、モデルが採点を意識して振る舞いを変える「評価認識」という最も懸念すべき兆候も自ら開示している。

Googleは2026年2月19日、同社のフラッグシップAIモデルの最新版「Gemini 3.1 Pro」をプレビュー公開した。2025年11月のGemini 3リリースからわずか数ヶ月という異例のスピードで開発・投入さ […]

2026年1月22日(木)、米国株式市場における今年最初の主要な暗号資産(仮想通貨)関連企業のIPOとして、デジタル資産カストディ(保管・管理)大手BitGo(ティッカー:BTGO)がニューヨーク証券取引所(NYSE)に […]

AIチャットボット「Claude」で知られるAI企業のAnthropicが、新たに最大20億ドル(約3000億円)の資金調達を計画していることが複数の報道により明らかになった。この調達が実現すれば、同社の企業価値は600 […]

Metaは、次世代大規模言語モデル(LLM)である「Llama 3」をリリースした。同社によれば、現在リリースされているほとんどのAIモデルよりも優れた性能を発揮するとしており、近いうちにマルチモダリティとより多くの言語 […]
Today, upgrading Apache Spark versions typically involves significant effort, with unclear investment requirements, including trial and error. This is mainly because in Spark there is no clear separation between the application code and the engine code. Apart from changes in the public API, any Spark internal changes may affect user workloads as users may rely on Spark internals: bug fixes in the Spark engine, changes to internal APIs, library upgrades, or language upgrades may affect customer workloads. As a result, Spark users are often reluctant to upgrade. The downside is that performance improvements, bug fixes, and new features take significantly longer to adopt, preventing customers from quickly benefiting from these improvements. In addition, it increases engineering complexity to manage a large number of different Spark versions. For Databricks serverless jobs and notebooks, we fundamentally transformed and simplified the user experience when using Spark. We shifted user focus from managing Spark runtime versions to managing the stable API that they integrate with - we created client-versioned workloads with a versionless Spark server. Decoupling the client from the Spark engine using Spark Connect has enabled Databricks to automatically upgrade the Spark server, providing users faster access to the latest features while minimizing disruptions from both intentional and unintentional breaking changes, all without compromising workload compatibility and with zero code changes needed from the user. This approach also offers significant benefits to Databricks by streamlining the release process, consolidating usage onto fewer server versions, and reducing engineering overhead from needing to backport changes. In this demonstration, we will first briefly introduce the architectural foundation of versionless Spark, leveraging Databricks' multi-user Spark compute and Spark Connect, followed by describing in more detail how we manage seamless upgrades for our customers, and finally talk about what the user experience is.
This paper provides a system literature review of the implementation of Generative Artificial Intelligence (GenAI) in ELT (Extract, Load, Transform) pipelines to incoming applications, concentrating on the Databricks and Snowflake services. The review is based on the summary of the results of fifty chosen studies devoted to the study of GenAI-based automation, scalability and adaptive transformation in real-time data processing. It is shown that GenAI drastically increases the intelligence of the pipeline and its work efficiency and allows working with dynamic schemas and with customised analytics. Nevertheless, data quality, data governance, explainability, and human control are still largely on the agenda. The research suggests a pathway to hybrid ELT architectures to combine GenAI automation and sound governance procedures to establish reliability and responsible execution in the streaming setting.
Managing and analyzing data in data lakes for big data environments requires robust protocols to ensure security, scalability, and compliance with privacy regulations. The increasing need to process sensitive data emphasizes the relevance of secure-by-design approaches that integrate encryption techniques and governance frameworks to protect personal and confidential information. This study proposes a protocol that combines the capabilities of Databricks and format-preserving encryption to improve data security and accessibility in data lakes without compromising usability or structure. The protocol was developed using a design science methodology, incorporating findings from a systematic literature review and validated through expert feedback and proof-of-concept experiments in banking environments. The proposed solution integrates multiple layers, data ingestion, persistence, access, and consumption, leveraging the processing capabilities of Databricks and format-preserving encryption to enable secure data management and governance. Validation results indicate the protocol is effectiveness in protecting sensitive data, with promising applicability in regulated industries. This work contributes to addressing key challenges in big data security and lays the groundwork for future developments in data governance and encryption techniques.
AI-augmented real-time retail analytics represents a transformative approach for modern retail operations, enabling businesses to process and act on data instantaneously in an increasingly competitive landscape. This comprehensive technical article explores the architecture, implementation, and business applications of an integrated analytics platform built on Apache Spark, Databricks, and Azure Event Hubs. The platform ingests data from diverse sources including IoT devices, point-of-sale systems, e-commerce platforms, mobile applications, and social media to create a unified view of retail operations. Advanced machine learning capabilities enable demand forecasting, customer segmentation, price optimization, and fraud detection with unprecedented accuracy. Large language models further enhance the platform by enabling natural language queries and automated insight generation, democratizing access to analytics across retail organizations. The business impact encompasses hyper-personalized customer experiences, predictive inventory management, revenue optimization strategies, and operational efficiency improvements. Implementation considerations and future trends are discussed, providing a blueprint for retailers seeking to leverage real-time analytics as a competitive differentiator in the age of artificial intelligence.