Google Launches New Search Gesture Feature for iPhone Users

2025-02-20

Google introduced Lens Screen Search for iPhone users via the Google app and Chrome browser, allowing quick searches by selecting screen content with gestures. This feature resembles Android's "circle search" but is limited to Google apps on iOS. Users can search text, images, or videos without leaving the app and may see AI-generated summaries in results. Future updates will include a Lens icon in the address bar for easier access.

Read more

Microsoft Launches Muse AI Model, Ushering in a New Era for Game Development

2025-02-20

Microsoft unveiled Muse, an AI model aiding Xbox game development by generating game environments, understanding 3D worlds, and responding to player interactions. Trained on data from "Bleeding Edge," Muse processes visual input but currently outputs low-resolution graphics. It supports early game development, enhances classic games, and aids in prototyping. Xbox CEO Phil Spencer envisions AI enabling old games to run on modern platforms without original hardware. While exploring AI's potential,

Read more

Former OpenAI CTO Mira Murati Launches Mind Machines Lab

2025-02-19

Mira Murati, ex-CTO of OpenAI, launches Mind Machines Lab, an AI startup focusing on multimodal models. The team includes former OpenAI leaders and aims to develop customizable, open-source AI technology while prioritizing infrastructure quality and safety. Initially backed by over $100 million in funding discussions, the company plans to advance research through community collaboration and publications. With 29 employees, it targets broad human expertise applications beyond programming tasks.

Read more

Meta to Host LlamaCon Conference Focusing on Open-Source AI Advances

2025-02-19

Meta will host LlamaCon on April 29, 2025, focusing on open-source AI development for developers. The event precedes Meta Connect in September, which highlights Meta Horizon advancements. LlamaCon's name derives from Meta's Llama AI model, with more details to come. Meta plans significant AI investments and may release new smart glasses. Ray-Ban Smart Glasses, launched in 2023, have sold over two million units.

Read more

Google's AI Assistant Gemini May Add Video Generation Feature

2025-02-19

Google's AI assistant Gemini may soon gain video generation capabilities, as indicated by code descriptions found in the Google app's APK file. The term "videogen" and Gemini's internal code name "robin" suggest an upcoming feature allowing users to create videos via simple AI commands. Although still under development, this advancement could revolutionize video creation while raising concerns over copyright and privacy. Google has yet to officially confirm the update.

Read more

Humane Sells Majority of Its Business to HP; AI Pin Product to Discontinue Service

2025-02-19

HP acquires majority of Humane's business for $116 million, including its CosmOS system, tech team, and over 300 patents. Humane will discontinue the AI Pin, with server support ending on February 28, 2025, impacting cloud-dependent features. Refunds are limited to recent purchases or prorated subscriptions. Initially priced at $1 billion, weak sales led to the lower acquisition. Humane's founders will lead HP's new AI division, HP IQ, integrating AI into PCs, printers, and meeting solutions.

Read more

Gemini Deep Research Features Now Available on Mobile Devices

2025-02-19

Gemini's Deep Research feature is now available on mobile devices, enhancing convenience for Gemini Advanced users. Originally launched for web, it provides comprehensive market analysis and aids investment decisions. Users can access detailed financial insights on Android or iOS anytime, anywhere, with web-based functionality still fully supported.

Read more

Google Meet Upgrade: AI-Generated Meeting Action Items

2025-02-19

Google Meet enhances its Gemini-powered note-taking tool with a new action item feature, generating next steps, assigning tasks, and setting deadlines post-meeting. Building on last year's voice-to-text functionality, this update aims to improve team collaboration efficiency through AI-driven automation. The rollout begins today but will proceed slower than usual to ensure quality.

Read more

Elon Musk Reveals Training Cost Data for Grok 3 for the First Time

2025-02-18

Musk unveiled Grok 3, an AI model with a training cost equivalent to 200,000 NVIDIA GPUs. Trained at xAI's advanced Colossus data center, Grok 3's scale is ten times larger than Grok 2. This significant investment in resources could lead to major advancements in reasoning, comprehension, and content generation, potentially reshaping the AI landscape.

Read more

Moon's Dark Side Open Platform Launches New Model kimi-latest to Meet Diversified Needs

2025-02-18

Dark Moon Corporation launched Kimi-Latest, a new visual model in the Kimi assistant line. It features advanced context processing up to 128k tokens, automatic model selection for billing, image understanding, and cost-saving caching mechanisms. Retaining functionalities from moonshot-v1, it suits chat apps and AI assistants but may require prompt revisions. For intent recognition tasks, moonshot-v1 is preferable. Note: Kimi-Latest only supports standard Kimi models; API access to Kimi k1.5 requ

Read more

Grok 3 Can Create 3D Models to Simulate Space Travel

2025-02-18

Grok 3 developed a 3D model visualizing Earth-to-Mars space travel, featuring accurate planetary representations and a simulated spacecraft journey. The model also demonstrates the return trip, showcasing its reliability. SpaceX plans to send Optimus and Grok to Mars via Starship within two years, advancing robotics in space exploration.

Read more

xAI Releases Latest Large Model Grok 3 with Significant Performance Improvement

2025-02-18

Elon Musk's xAI launched Grok 3, an advanced family of large language models surpassing competitors in math, science, and code writing. Key features include DeepSearch for summarized web queries, optimized reasoning variants, and a subscription service called SuperGrok. Developed using Colossus supercomputers with 100,000 NVIDIA H100 GPUs, Grok 3 offers faster performance and improved accuracy. Voice mode is delayed due to unresolved issues, while Grok 2 will be open-sourced after Grok 3 stabili

Read more

Step Stars Collaborates with Geely to Open Source the Step Series Multimodal Large Models for Video and Audio Fields

2025-02-18

Stairway Stars and Geely Automotive Group announced the global open-source release of their co-developed Step series multimodal models. The series includes Step-Video-T2V, a 30-billion-parameter video generation model, and Step-Audio, an industry-first open-source voice interaction model. Step-Video-T2V generates high-quality 540P videos, while Step-Audio supports emotional, personalized voice interactions with natural expression. Users can explore Step-Audio via the Yuewen App, advancing AI inn

Read more

Kunlun Tech Open-Sources SkyReels-V1, China's First Video Generation Model for AI-Powered Short Drama Creation

2025-02-18

Kunlun Tech introduced two AI video technologies: SkyReels-V1, an open-source model for realistic short drama creation with detailed human performances, and SkyReels-A1, an algorithm enhancing facial and motion control. Both aim to reduce costs, increase accessibility, and improve efficiency in AI-driven video generation, enabling broader applications in content creation.

Read more

DarkMind: A New Backdoor Attack Leveraging LLM Inference Capabilities

2025-02-18

Researchers from the University of St. Louis developed DarkMind, a stealthy backdoor attack targeting large language models (LLMs). It manipulates the reasoning process without detection, activating during specific steps to alter outputs. Unlike traditional attacks, DarkMind embeds hidden triggers within custom LLMs, remaining dormant until triggered. This vulnerability affects advanced models like GPT-4o and LLaMA-3, posing risks to security and reliability across various domains.

Read more

Mistral Releases Custom Language Model for the Arab Region

2025-02-18

Paris-based AI startup Mistral unveiled Mistral Saba, a specialized model for Arabic-speaking countries with 2.4 billion parameters. It outperforms Mistral's general-purpose models in Arabic and Indo-Aryan languages. Saba aims to enhance conversational support and content generation while targeting Middle Eastern clients. Accessible via API or internal deployment, it caters to sectors like finance and healthcare. This launch expands Mistral's multilingual focus and strategic presence in the Midd

Read more

Apple's Cartoon Avatar Generator Exposes Bias Issues

2025-02-18

Apple's Image Playground app faces criticism for racial bias, struggling to recognize a machine learning expert's skin tone and hair texture. Despite limiting functionality to cartoon-style images, biases persist. This highlights ongoing concerns about fairness and accuracy in AI image generation as a critical issue requiring attention.

Read more

Musk's xAI Company Set to Launch Grok 3 Chatbot

2025-02-17

Elon Musk's xAI to unveil Grok 3, billed as the "smartest AI on Earth," featuring advanced reasoning and multimodal capabilities. Integrated into X platform, Grok 3 follows Grok 2, which improved real-world performance. Competing with OpenAI's GPT 4o upgrades, xAI negotiates $5 billion in server deals with Dell and raises $10 billion for GPU investments, though its $51 billion valuation lags behind OpenAI's $150 billion.

Read more

Baidu Search Fully Integrates DeepSeek and Wenxin Large Model Deep Search Technology

2025-02-17

Baidu Search announces an update integrating DeepSeek and Wenxin Large Model for enhanced search experiences. These features, available to all users, offer accurate information retrieval and support multi-modal input/output. Developers on the Wenxin platform can leverage DeepSeek to optimize products, boosting ecosystem growth. This integration promises improved efficiency and innovation for both users and developers.

Read more