News Gist .News

Articles | Politics | Finance | Stocks | Crypto | AI | Technology | Science | Gaming | PC Hardware | Laptops | Smartphones | Archive

Super Mario to Benchmark AI Performance.

Researchers at Hao AI Lab have used Super Mario Bros. as a benchmark for AI performance, with Anthropic's Claude 3.7 performing the best, followed by Claude 3.5. This unexpected choice highlights the limitations of traditional benchmarks in evaluating AI capabilities. The lab's approach demonstrates the need for more nuanced and realistic evaluation methods to assess AI intelligence.

See Also

AI Versus the Brain and the Race for General Intelligence Δ1.78

The ongoing debate about artificial general intelligence (AGI) emphasizes the stark differences between AI systems and the human brain, which serves as the only existing example of general intelligence. Current AI, while capable of impressive feats, lacks the generalizability, memory integration, and modular functionality that characterize brain operations. This raises important questions about the potential pathways to achieving AGI, as the methods employed by AI diverge significantly from those of biological intelligence.

AI Scholars Win Turing Prize for Technique That Made Possible AlphaGo's Chess Triumph Δ1.75

Andrew G. Barto and Richard S. Sutton have been awarded the 2025 Turing Award for their pioneering work in reinforcement learning, a key technique that has enabled significant achievements in artificial intelligence, including Google's AlphaZero. This method operates by allowing computers to learn through trial and error, forming strategies based on feedback from their actions, which has profound implications for the development of intelligent systems. Their contributions not only laid the mathematical foundations for reinforcement learning but also sparked discussions on its potential role in understanding creativity and intelligence in both machines and living beings.

The Unstoppable Artificial Intelligence (AI) Stock That Could Join the $3 Trillion Club by 2028 Δ1.75

Meta Platforms is poised to join the exclusive $3 trillion club thanks to its significant investments in artificial intelligence, which are already yielding impressive financial results. The company's AI-driven advancements have improved content recommendations on Facebook and Instagram, increasing user engagement and ad impressions. Furthermore, Meta's AI tools have made it easier for marketers to create more effective ads, leading to increased ad prices and sales.

AI Bots Can Now Play Mafia with Each Other, and Almost All of Them Are Terrible at It Δ1.74

The AI Language Learning Models (LLMs) playing Mafia with each other have been entertaining, if not particularly skilled. Despite their limitations, the models' social interactions and mistakes offer a glimpse into their capabilities and shortcomings. The current LLMs struggle to understand roles, make alliances, and even deceive one another. However, some models, like Claude 3.7 Sonnet, stand out as exceptional performers in the game.

AI Startup Anthropic Valued at $61.5B After Latest Funding Round. Δ1.74

Anthropic has secured a significant influx of capital, with its latest funding round valuing the company at $61.5 billion post-money. The Amazon- and Google-backed AI startup plans to use this investment to advance its next-generation AI systems, expand its compute capacity, and accelerate international expansion. Anthropic's recent announcements, including Claude 3.7 Sonnet and Claude Code, demonstrate its commitment to developing AI technologies that can augment human capabilities.

Microsoft Accelerates AI Efforts to Compete with OpenAI Δ1.74

In accelerating its push to compete with OpenAI, Microsoft is developing powerful AI models and exploring alternatives to power products like Copilot bot. The company has developed AI "reasoning" models comparable to those offered by OpenAI and is reportedly considering offering them through an API later this year. Meanwhile, Microsoft is testing alternative AI models from various firms as possible replacements for OpenAI technology in Copilot.

AI Takes Center Stage as Alibaba Drives Shares Higher Δ1.74

Alibaba Group's release of an artificial intelligence (AI) reasoning model has driven its Hong Kong-listed shares more than 8% higher on Thursday, outperforming global hit DeepSeek's R1. The company's AI unit claims that its QwQ-32B model can achieve performance comparable to top models like OpenAI's o1 mini and DeepSeek's R1. Alibaba's new model is accessible via its chatbot service, Qwen Chat, allowing users to choose various Qwen models.

The Ai Arms Race Heats Up: Tencent Unveils Model that Outdoes Deepseek Δ1.74

Tencent Holdings Ltd. has unveiled its Hunyuan Turbo S artificial intelligence model, which the company claims outperforms DeepSeek's R1 in response speed and deployment cost. This latest move joins a series of rapid rollouts from major industry players on both sides of the Pacific since DeepSeek stunned Silicon Valley with a model that matched the best from OpenAI and Meta Platforms Inc. The Hunyuan Turbo S model is designed to respond as instantly as possible, distinguishing itself from the deep reasoning approach of DeepSeek's eponymous chatbot.

Compare AI Models Simplifies Evaluation of AI Technologies Δ1.74

Compare AI Models is an online platform that facilitates the assessment and comparison of various AI models using key performance indicators. It caters to businesses, developers, and researchers by providing structured comparisons across over 20 large language models and other AI technologies, thereby streamlining the decision-making process. While the tool offers valuable insights into model capabilities, it does not generate content or allow for fine-tuning, making it essential for users to understand its limitations.

Conan O'Brien Comments on AI During Oscars Opening Monologue Δ1.73

When hosting the 2025 Oscars last night, comedian and late-night TV host Conan O’Brien addressed the use of AI in his opening monologue, reflecting the growing conversation about the technology’s influence in Hollywood. Conan jokingly stated that AI was not used to make the show, but this remark has sparked renewed debate about the role of AI in filmmaking. The use of AI in several Oscar-winning films, including "The Brutalist," has ignited controversy and raised questions about its impact on jobs and artistic integrity.

The Ai Chatbot App Gains Global Momentum as Deepseek Surpasses U.s. Competition Δ1.73

DeepSeek has broken into the mainstream consciousness after its chatbot app rose to the top of the Apple App Store charts (and Google Play, as well). DeepSeek's AI models, trained using compute-efficient techniques, have led Wall Street analysts — and technologists — to question whether the U.S. can maintain its lead in the AI race and whether the demand for AI chips will sustain. The company's ability to offer a general-purpose text- and image-analyzing system at a lower cost than comparable models has forced domestic competition to cut prices, making some models completely free.

3 Best Artificial Intelligence (AI) Stocks to Buy in March Δ1.73

Amid recent volatility in the AI sector, investors are presented with promising opportunities, particularly in stocks like Nvidia, Amazon, and Microsoft. Nvidia, despite a notable decline from its peak, continues to dominate the GPU market, essential for AI development, while Amazon's cloud computing division is significantly investing in AI infrastructure. The current market conditions may favor long-term investors who strategically identify undervalued stocks with substantial growth potential in the burgeoning AI industry.

New Ai Text Diffusion Models Break Speed Barriers by Pulling Words From Noise Δ1.73

These diffusion models maintain performance faster than or comparable to similarly sized conventional models. LLaDA's researchers report their 8 billion parameter model performs similarly to LLaMA3 8B across various benchmarks, with competitive results on tasks like MMLU, ARC, and GSM8K. Mercury claims dramatic speed improvements, operating at 1,109 tokens per second compared to GPT-4o Mini's 59 tokens per second.

Openai’s Largest Ai Model Ever Arrives to Mixed Reviews Δ1.73

GPT-4.5 offers marginal gains in capability but poor coding performance despite being 30 times more expensive than GPT-4o. The model's high price and limited value are likely due to OpenAI's decision to shift focus from traditional LLMs to simulated reasoning models like o3. While this move may mark the end of an era for unsupervised learning approaches, it also opens up new opportunities for innovation in AI.

Microsoft Pushes Ahead with Ai in Gaming Δ1.73

Microsoft is exploring the potential of AI in its gaming efforts, as revealed by the Muse project, which can generate gameplay and understand 3D worlds and physics. The company's use of AI has sparked debate among developers, who are concerned that it may replace human creators or alter the game development process. Microsoft's approach to AI in gaming is seen as a significant step forward for the industry.

Pioneers of Reinforcement Learning Win the Turing Award Δ1.73

The 2023 Turing Award winners, Andrew Barto and Rich Sutton, have been recognized for their work in reinforcement learning, a crucial component of artificial intelligence that enables machines to learn from experience. Their research has led to significant advancements in machine learning, paving the way for applications in robotics, game playing, and more. The award acknowledges the pioneers' contributions to this rapidly evolving field.

Barclays: Super Micro's Ai Lead Shrinking, Margin Targets at Risk Δ1.73

Super Micro faces uncertainty in AI server demand, as Barclays highlights margin pressures and a shrinking competitive moat. The company's reliance on Nvidia's Blackwell products has raised concerns about its ability to maintain its market share. Despite its leadership in AI servers, Super Micro is facing significant challenges, including limited visibility on build orders and steep learning curves.

Hugging Face's Chief Science Officer Worries AI Is Becoming 'Yes-Men on Servers' Δ1.73

Thomas Wolf, co-founder and chief science officer of Hugging Face, expresses concern that current AI technology lacks the ability to generate novel solutions, functioning instead as obedient systems that merely provide answers based on existing knowledge. He argues that true scientific innovation requires AI that can ask challenging questions and connect disparate facts, rather than just filling in gaps in human understanding. Wolf calls for a shift in how AI is evaluated, advocating for metrics that assess the ability of AI to propose unconventional ideas and drive new research directions.

Amazon Is Reportedly Developing Its Own AI 'Reasoning' Model Δ1.73

Amazon is reportedly venturing into the development of an AI model that emphasizes advanced reasoning capabilities, aiming to compete with existing models from OpenAI and DeepSeek. Set to launch under the Nova brand as early as June, this model seeks to combine quick responses with more complex reasoning, enhancing reliability in fields like mathematics and science. The company's ambition to create a cost-effective alternative to competitors could reshape market dynamics in the AI industry.

The AI Chatbot Showdown Reveals No Clear Winner Δ1.73

GPT-4.5 and Google's Gemini Flash 2.0, two of the latest entrants to the conversational AI market, have been put through their paces to see how they compare. While both models offer some similarities in terms of performance, GPT-4.5 emerged as the stronger performer with its ability to provide more detailed and nuanced responses. Gemini Flash 2.0, on the other hand, excelled in its translation capabilities, providing accurate translations across multiple languages.

Detecting Deception in Digital Content Δ1.72

SurgeGraph has introduced its AI Detector tool to differentiate between human-written and AI-generated content, providing a clear breakdown of results at no cost. The AI Detector leverages advanced technologies like NLP, deep learning, neural networks, and large language models to assess linguistic patterns with reported accuracy rates of 95%. This innovation has significant implications for the content creation industry, where authenticity and quality are increasingly crucial.

Openai Unveils gpt-4.5 'Orion,' Its Largest Ai Model Yet Δ1.72

OpenAI has launched GPT-4.5, a significant advancement in its AI models, offering greater computational power and data integration than previous iterations. Despite its enhanced capabilities, GPT-4.5 does not achieve the anticipated performance leaps seen in earlier models, particularly when compared to emerging AI reasoning models from competitors. The model's introduction reflects a critical moment in AI development, where the limitations of traditional training methods are becoming apparent, prompting a shift towards more complex reasoning approaches.

AI Stocks on Hedge Funds' Radar: A Closer Look at Alibaba Group Holding Limited (BABA) Δ1.72

Alibaba Group Holding Limited (NYSE:BABA) stands out among AI stocks as a leader in the field of artificial intelligence, with significant investments and advancements in its latest GPT-4.5 model. The company's enhanced ability to recognize patterns, generate creative insights, and show emotional intelligence sets it apart from other models. Early testing has shown promising results, with the model hallucinating less than others.

Monster Hunter Wilds Benchmark - Demanding Action Role-Playing Game Hit Needs a dGPU Δ1.72

The Monster Hunter Wilds Benchmark reveals that the game has high hardware demands, with even powerful integrated graphics cards struggling to maintain smooth performance. Testing across various resolutions and settings indicates that a mid-range graphics card is necessary for optimal gameplay, especially in demanding combat scenarios. The benchmark results highlight the importance of upscaling technologies like DLSS and FSR for achieving playable frame rates on less powerful systems.