Loading video...

Video Failed to Load

Go Home

Skala 1.1, the updated deep-learning exchange-correlation functional from Microsoft Research, provides greater accuracy, expanded accessibility across the computational chemistry ecosystem, and a living benchmark to track computational performance.

18,451 views • 22 days ago •via X (Twitter)

6 Comments

Love Web3 World's profile picture
Love Web3 World21 days ago

2.8 kcal/mol on GMTKN55 is a serious result but accessibility may matter even more. Skala 1.1 brings deep-learning DFT closer to everyday computational chemistry, with CP2K support available and broader integrations underway.

Mohammed Benaissa's profile picture
Mohammed Benaissa21 days ago

Can it predict good band gaps

Zohaib Ai's profile picture
Zohaib Ai21 days ago

Great advancement! 🧪 Skala 1.1 pushes computational chemistry forward.

Farrukh Masud's profile picture
Farrukh Masud21 days ago

Accessibility might end up mattering more than the accuracy gains. Getting these methods into the tools people already use (instead of staying stuck in research papers) is usually what determines whether something actually gets adopted. The living benchmark is a smart addition too.

Kelvin Tran's profile picture
Kelvin Tran21 days ago

Nghe hấp dẫn ghê, không biết bản này có chạy nhanh hơn bản cũ không nhỉ? 🧪

Fajar M Reza's profile picture
Fajar M Reza21 days ago

Living benchmarks matter when scientific models must improve beyond one static score.

Related Videos

Computing hardware architecture has evolved from maximizing single-core performance to multiplying parallel processing capacity through more cores, more threads, and higher computational density. This transition demands software architectures specifically designed to exploit parallel processing, especially in advanced AI and blockchain-based systems. As Greg Meredith, CEO of our partner F1R3FLY. io, notes, "We need software architectures that can eat physical threads of computation and turn that into throughput." Traditional sequential programming models often struggle to effectively leverage the potential of these multi-core systems, creating a bottleneck in performance scaling. Rholang addresses this challenge through its foundation in the Rho-Calculus, a reflective higher-order process calculus specifically designed for inherent concurrency and reflection. Unlike computational models based on sequential execution or those that bolt on concurrency primitives, Rholang can leverage its process-oriented nature to detect and distribute non-interfering logical computation threads across available physical processing units, creating a more direct correlation between hardware resources and performance scaling. The Rho-Calculus was specifically designed to provide the minimal set of operations needed to model autonomous agents operating in parallel while maintaining the ability to represent their own reasoning processes. This structure mirrors the operational characteristics of intelligent systems, where autonomous processes run independently while communicating and coordinating. For our novel decentralized AI platform, MeTTaCycle, this future-proof architectural foundation enables more than just enhanced transactional throughput; it fosters an environment where independent computational processes can efficiently coexist and interact, forming a critical layer for the Artificial Superintelligence Alliance's decentralized AGI infrastructure and products.

SingularityNET

29,368 views • 1 year ago

We’re launching Optima. Now anyone can create a custom benchmark for their use case, leveraging Artificial Analysis’ leading research and platform Building and running benchmarks is difficult. We have distilled Artificial Analysis’ research and experience developing benchmarks into Optima, a new platform for benchmarking models on your own workloads and comparing performance, speed and cost efficiency. Optima allows you to find the best model for your task, or an equally performant alternative to your current setup at 10x lower cost or time per task. We’ve integrated Artificial Analysis' research and experience in benchmarks across the Optima workflow: ➤ Build benchmarks based on your own data and use cases: There are three ways to build a benchmark with Optima. Upload an existing evaluation dataset from your own files or Hugging Face, or import agent traces from platforms including Arize AI, Braintrust and langfuse.com. Install the Optima skill to build a benchmark using context from your coding environment and previous sessions. Or simply describe your use case and provide example inputs and outputs, and Optima will build the benchmark for you ➤ Run across the latest models: Run the same benchmark across leading models in a single click, and keep your leaderboard up to date as soon as new models are released ➤ Bring Artificial Analysis grading to your own benchmark: Evaluate responses against objective rubric criteria or using the same pairwise judging approach used for Artificial Analysis benchmarks including GDPval-AA and AA-Briefcase. For pairwise judging, select your preferred responses from a sample and Optima uses those preferences to rank models across your test set ➤ Compare performance, cost and time efficiency: Optima measures more than model performance. Cost per Task and Time per Task are tracked alongside benchmark scores, with category-level results and support for custom metrics, allowing you to compare the tradeoffs between models for your specific use case Ahead of launch, here are examples questions our beta testers answered with Optima: ➤ Which model can save me 10x the cost without a meaningful decrease in quality for my finance & accounting agent? ➤ Which model best matches the writing style of lawyers for my legal agent? ➤ Which model can best identify different elements in my custom image dataset? Optima is available today. Build your own benchmark at

Artificial Analysis

132,623 views • 1 month ago