Connect with us

Technology

S&P Global Introduces a Novel Artificial Intelligence Benchmark for the Financial Sector

Published

on

The debut of S&P AI Benchmarks by Kensho was discreetly announced by S&P Global, a prominent supplier of financial intelligence, on Wednesday. Large language models (LLMs) are used in complex financial and quantitative applications, and this creative solution seeks to establish a new benchmark for assessing LLM performance.

Created by Kensho, S&P Global’s AI-focused division, the benchmarking tool evaluates an LLM’s proficiency in tasks like quantitative reasoning, extracting data from financial documents, and proving domain-specific knowledge. Each model’s capabilities are transparently shown after the results are shown on a leaderboard.

According to Bhavesh Dayalji, CEO of Kensho and Chief AI Officer for S&P Global, “S&P AI Benchmarks combined Kensho’s cutting-edge AI research and engineering with S&P Global’s leading financial intelligence capabilities” in a VentureBeat interview. “We hope that the solution spurs more innovation in the FinAI space and becomes the industry standard for understanding how LLMs perform on complex financial reasoning.”

As more financial services organizations investigate the possibilities of generative AI and LLMs to improve efficiency and gain a competitive edge, the introduction of S&P AI Benchmarks is timely. Unfortunately, organizations find it difficult to evaluate which models are best suited for their particular use cases due to the absence of established benchmarks.

Encouraging Creativity and Thoughtful Judgment

According to Dayalji, “benchmark solutions like ours are critical to helping institutions and professionals across our industry determine which LLMs they should be using for their specific use cases.” In addition, “we think that S&P AI Benchmarks will spur innovation by assisting financial professionals in determining where each model excels and where it can bring the greatest benefit.”

Engineers, researchers, academics, and financial professionals from all of S&P Global’s divisions made up the diverse team of experts who developed and validated the S&P AI Benchmarks methodology. An LLM’s performance in three important categories is rigorously tested by the 600 questions that make up the evaluation set.

A Milestone for AI Adoption in Finance

The introduction of S&P AI Benchmarks, according to industry analysts, may represent a critical turning point in the financial sector’s adoption of AI. With increasingly sophisticated AI influencing the financial sector, companies will need a trustworthy and transparent benchmarking tool to help them choose which models to use. The FinAI space may see innovation spurred by S&P Global’s solution, which could also hasten the responsible adoption of LLMs.

S&P Global believes that S&P AI Benchmarks will be instrumental in influencing how AI will develop in the financial services industry going forward. “Our goal is for LLMs to become more efficient and better tailored to the demands of all the industries we serve, and our solutions will help us get there,” Dayalji stated. “In order for us to keep improving our framework, we encourage participation from all model providers.”

Tools like S&P AI Benchmarks by Kensho, which help organizations harness the power of AI and generative AI while ensuring accuracy, transparency, and responsible deployment, are poised to become indispensable guides as the financial industry navigates the rapidly evolving landscape of these technologies.

Technology

Microsoft Expands Copilot Voice and Think Deeper

Published

on

Microsoft Expands Copilot Voice and Think Deeper

Microsoft is taking a major step forward by offering unlimited access to Copilot Voice and Think Deeper, marking two years since the AI-powered Copilot was first integrated into Bing search. This update comes shortly after the tech giant revamped its Copilot Pro subscription and bundled advanced AI features into Microsoft 365.

What’s Changing?

Microsoft remains committed to its $20 per month Copilot Pro plan, ensuring that subscribers continue to enjoy premium benefits. According to the company, Copilot Pro users will receive:

  • Preferred access to the latest AI models during peak hours.
  • Early access to experimental AI features, with more updates expected soon.
  • Extended use of Copilot within popular Microsoft 365 apps like Word, Excel, and PowerPoint.

The Impact on Users

This move signals Microsoft’s dedication to enhancing AI-driven productivity tools. By expanding access to Copilot’s powerful features, users can expect improved efficiency, smarter assistance, and seamless integration across Microsoft’s ecosystem.

As AI technology continues to evolve, Microsoft is positioning itself at the forefront of innovation, ensuring both casual users and professionals can leverage the best AI tools available.

Stay tuned for further updates as Microsoft rolls out more enhancements to its AI offerings.

Continue Reading

Technology

Google Launches Free AI Coding Tool for Individual Developers

Published

on

Google Launches Free AI Coding Tool for Individual Developers

Google has introduced a free version of Gemini Code Assistant, its AI-powered coding assistant, for solo developers worldwide. The tool, previously available only to enterprise users, is now in public preview, making advanced AI-assisted coding accessible to students, freelancers, hobbyists, and startups.

More Features, Fewer Limits

Unlike competing tools such as GitHub Copilot, which limits free users to 2,000 code completions per month, Google is offering up to 180,000 code completions—a significantly higher cap designed to accommodate even the most active developers.

“Now anyone can easily learn, generate code snippets, debug, and modify applications without switching between multiple windows,” said Ryan J. Salva, Google’s senior director of product management.

AI-Powered Coding Assistance

Gemini Code Assist for individuals is powered by Google’s Gemini 2.0 AI model and offers:
Auto-completion of code while typing
Generation of entire code blocks based on prompts
Debugging assistance via an interactive chatbot

The tool integrates with popular developer environments like Visual Studio Code, GitHub, and JetBrains, supporting a wide range of programming languages. Developers can use natural language prompts, such as:
Create an HTML form with fields for name, email, and message, plus a submit button.”

With support for 38 programming languages and a 128,000-token memory for processing complex prompts, Gemini Code Assist provides a robust AI-driven coding experience.

Enterprise Features Still Require a Subscription

While the free tier is generous, advanced features like productivity analytics, Google Cloud integrations, and custom AI tuning remain exclusive to paid Standard and Enterprise plans.

With this move, Google aims to compete more aggressively in the AI coding assistant market, offering developers a powerful and unrestricted alternative to existing tools.

Continue Reading

Technology

Elon Musk Unveils Grok-3: A Game-Changing AI Chatbot to Rival ChatGPT

Published

on

Elon Musk Unveils Grok-3: A Game-Changing AI Chatbot to Rival ChatGPT

Elon Musk’s artificial intelligence company xAI has unveiled its latest chatbot, Grok-3, which aims to compete with leading AI models such as OpenAI’s ChatGPT and China’s DeepSeek. Grok-3 is now available to Premium+ subscribers on Musk’s social media platform x (formerly Twitter) and is also available through xAI’s mobile app and the new SuperGrok subscription tier on Grok.com.

Advanced capabilities and performance

Grok-3 has ten times the computing power of its predecessor, Grok-2. Initial tests show that Grok-3 outperforms models from OpenAI, Google, and DeepSeek, particularly in areas such as math, science, and coding. The chatbot features advanced reasoning features capable of decomposing complex questions into manageable tasks. Users can interact with Grok-3 in two different ways: “Think,” which performs step-by-step reasoning, and “Big Brain,” which is designed for more difficult tasks.

Strategic Investments and Infrastructure

To support the development of Grok-3, xAI has made major investments in its supercomputer cluster, Colossus, which is currently the largest globally. This infrastructure underscores the company’s commitment to advancing AI technology and maintaining a competitive edge in the industry.

New Offerings and Future Plans

Along with Grok-3, xAI has also introduced a logic-based chatbot called DeepSearch, designed to enhance research, brainstorming, and data analysis tasks. This tool aims to provide users with more insightful and relevant information. Looking to the future, xAI plans to release Grok-2 as an open-source model, encouraging community participation and further development. Additionally, upcoming improvements for Grok-3 include a synthesized voice feature, which aims to improve user interaction and accessibility.

Market position and competition

The launch of Grok-3 positions xAI as a major competitor in the AI ​​chatbot market, directly challenging established models from OpenAI and emerging competitors such as DeepSeek. While Grok-3’s performance claims are yet to be independently verified, early indications suggest it could have a significant impact on the AI ​​landscape. xAI is actively seeking $10 billion in investment from major companies, demonstrating its strong belief in their technological advancements and market potential.

Continue Reading

Trending

error: Content is protected !!