الوسم: Artificial Intelligence

  • ChatGPT Advanced Voice Feature Now Available for Plus and Teams

    ChatGPT Advanced Voice Feature Now Available for Plus and Teams

    The Advanced Voice Mode's UI
    OpenAI

    OpenAI took to Twitter on Tuesday to announce the rollout of its Advanced Voice feature alongside five new voices for its conversational AI. This feature will be available to Plus and Teams subscribers throughout the week, with Enterprise and Edu users set to gain access next week.

    The Advanced Voice feature operates on the GPT-4o model, allowing users to communicate with the chatbot using voice rather than typing. This capability was first introduced during OpenAI’s Spring Update event and underwent beta testing in July with a limited number of ChatGPT Plus subscribers. Now, it’s available to all paying subscribers for experimentation.

    Advanced Voice Mode's notification screen
    OpenAI

    In addition to Advanced Voice, OpenAI has introduced five new voices for the chatbot: Arbor, Maple, Sol, Spruce, and Vale, which can be accessed in both Standard and Advanced Voice modes. These join the existing voices—Breeze, Juniper, Cove, and Ember. However, OpenAI noted that video and screen sharing will not be available in Advanced Voice mode just yet, with plans to include those features in future updates.

    Furthermore, OpenAI is integrating two tools to enhance Advanced Voice capabilities, aligning them more closely with its text-based experience: memory and custom instructions. Initially, Advanced Voice could only reference the current chat’s information. With the memory feature, the AI will be able to retrieve information from past conversations, reducing the need for users to repeat themselves. On the other hand, custom instructions allow users to set parameters for the model’s responses, such as specifying that coding responses be given in Python.

    Subscribers to Plus and Teams will receive an in-app notification alerting them when the feature is activated on their account. However, it’s important to note that Advanced Voice is not yet available in the EU, the U.K., Switzerland, Iceland, Norway, or Liechtenstein.

    OpenAI is not the only tech company enabling direct voice interactions with AI. This announcement from OpenAI follows closely on the heels of Google’s launch of its Gemini Live to all users, including those on the free tier.

  • Max-Planck Institute Unveils Innovative Modular Hexagon Design

    Max-Planck Institute Unveils Innovative Modular Hexagon Design

    Researchers at the Max-Planck-Institute for Intelligent Systems (MPI-IS) have made a groundbreaking advancement in robotic technology with the introduction of hexagon-shaped modular components, known as HEXEL modules. This innovative design allows for the rapid assembly and reconfiguration of high-speed robots, akin to the versatility of LEGO bricks. The findings, led by Christoph Keplinger and his team from the Robotic Materials Department, are set to be published in the prestigious journal Science Robotics on September 18, 2024.

    Each HEXEL module features a lightweight exoskeleton constructed from six rigid glass fiber plates, providing the necessary structure and durability. At the core of these modules are advanced artificial muscles known as hydraulically amplified self-healing electrostatic (HASEL) actuators. By applying high voltage, these artificial muscles activate and enable the hexagonal joints to transform from elongated shapes to broader, flatter forms.

    The unique combination of soft and rigid elements within the modules facilitates impressive movements and speeds. “By linking multiple modules together, we can generate new robot forms that can adapt to diverse operational requirements,” explains Ellen Rumley, a visiting researcher from the University of Colorado Boulder and co-first author of the upcoming publication.

    Demonstrations by the research team showcase the diverse capabilities of the HEXEL modules. In one instance, a collection of these modules maneuvers through a narrow passage, while a single unit performs rapid jumps into the air. Additionally, when configured into larger assemblies, the modules can achieve distinct motion patterns based on their arrangement—one such example being a robot capable of rolling at high speeds.

    Zachary Yoder, also a co-first author and Ph.D. student at MPI-IS, emphasizes the practicality of this modular approach. “Developing robots with adaptable designs is not only innovative but also sustainable. Instead of investing in multiple specialized robots for various tasks, users can construct a multitude of configurations using a common set of components. This flexibility can be especially advantageous in environments where resources are scarce,” he notes.

    The emergence of HEXEL modules represents a significant leap forward in robotics, promising enhanced functionality and efficiency while paving the way for future developments in modular design. As the research team continues to explore the potential of these innovative components, the horizons of robotics are broadening, ushering in a new era of adaptability and sophistication.

  • Copilot Wave 2: Discover New AI Features To Try Out

    Copilot Wave 2: Discover New AI Features To Try Out

    Microsoft has unveiled an exciting update to its AI assistant, Copilot, referred to as “Wave 2.” This update enhances Copilot’s functionality across various widely used Office applications and introduces features specifically designed for businesses, including the innovative Copilot Pages.

    Let’s dive into Copilot Pages first. Described as a “dynamic, persistent canvas,” Pages facilitate “multiplayer” collaboration seamlessly integrated within Copilot. Microsoft aims to create a more productive environment by allowing users to accomplish tasks directly within Copilot without switching between applications.

    Microsoft 365 Copilot | Copilot Pages

    Copilot Pages acts as a convenient hub where you can extract content from Copilot and store it in an easily accessible format. It encourages collaboration since multiple users can join and edit the same Page simultaneously. While it may seem similar to existing features found in other applications, Microsoft sees it as a means to deepen user engagement with Copilot, a key component of their strategy for AI integration.

    Microsoft succinctly explains: “Pages transform temporary AI-generated content into something lasting, allowing for editing, additions, and sharing with others.”

    This addition to Copilot comes at no extra cost, but access requires a Microsoft Entra account. The rollout is scheduled for the upcoming weeks, although details regarding wider accessibility were not specified.

    Along with Pages, Microsoft is also implementing updates across Word, Excel, PowerPoint, Outlook, Teams, and OneDrive. While the changes aren’t sweeping, they encapsulate a broad enhancement across the board:

    • Word: Users will be able to easily reference data from the web alongside information from other apps, featuring a new “on-canvas” experience that prominently integrates Copilot.
    • Excel: Python is now operable within Copilot, enabling advanced analysis through natural language commands and improved handling of unformatted data. Additionally, new skills are being implemented for more adept Excel users, with Python in public preview.
    • PowerPoint: A feature called Narrative Builder allows users to create presentations based solely on prompts while customizing outlines and adhering to corporate branding via approved SharePoint graphics.
    • Teams: Copilot will take on an enhanced role by analyzing meeting transcripts and chats, letting users ask questions about missed content.
    • Outlook: The “Prioritize my inbox” feature organizes emails using AI, focusing on key communications and ongoing threads.
    • OneDrive: An AI-powered search function enables Copilot to discern and retrieve information from files more efficiently.

    Lastly, Microsoft is introducing Copilot agents and a corresponding agent builder tool, found within Copilot Studio. While not widely available to the general public, these agents can be tailored for large organizations to automate various tasks, ranging from simple prompts to complex autonomous actions.

    The potential capabilities of these agents are intriguing, though most users may not interact with them directly. However, as developers and large companies begin utilizing these AI agents, they have the potential to significantly transform workflows.

  • OpenAI’s New ‘Project Strawberry’ Model Is Here

    OpenAI’s New ‘Project Strawberry’ Model Is Here


    chatGPT on a phone on an encyclopedia
    Shantanu Kumar / Pexels

    OpenAI has officially unveiled the production version of its groundbreaking reasoning model, previously known as Project Strawberry, now titled “o1.” Alongside it is a “mini” variant, similar to GPT-4o, which promises quicker, more efficient interactions at the cost of a more limited knowledge base.

    The o1 model showcases a variety of technical improvements. As the first model in OpenAI’s reasoning suite, o1 leverages human-like deduction to tackle complex inquiries spanning fields like science, coding, and mathematics, outperforming human speeds.

    During trials, o1 was tasked with a qualifying exam for the International Mathematics Olympiad. In a striking comparison, while its predecessor, GPT-4o, solved only 13% of the posed questions accurately, o1 achieved an impressive 83%. Furthermore, in an online Codeforces competition, o1 secured a spot in the 89th percentile. It has also demonstrated the ability to confidently handle questions that previously stumped earlier models, such as determining which number is larger, 9.11 or 9.9. However, OpenAI emphasizes that this is merely a glimpse of the model’s full potential.

    According to Jerry Tworek, the research lead at OpenAI, the new o1 “has been trained using a completely new optimization algorithm and a new training dataset specifically tailored for it.” By integrating reinforcement learning alongside a “chain of thought” approach, o1 is said to deliver more precise conclusions than prior iterations. Tworek noted that while the model has reduced instances of hallucination, “we can’t say we solved hallucinations.”

    Starting today, both ChatGPT-Plus and Teams subscribers can experience o1 and its mini counterpart. Enterprise and Edu subscribers can expect access within the upcoming week. Additionally, o1-mini will eventually be made available to free-tier users, although a specific timeline has not been disclosed.

    Developers should be aware that the API pricing for o1 represents a significant increase compared to GPT-4o. Accessing o1 will cost $15 for every million input tokens, in contrast to GPT-4o’s $5 rate, while output tokens will be priced at $60 per million tokens—fourfold higher than the previous model’s fee. This raises an intriguing question regarding o1’s ability to recognize the correct number of ‘R’s in the word “strawberry.”

  • Google’s Gemini Live Now Free on Android

    Google’s Gemini Live Now Free on Android


    Google has announced that its Gemini Live feature, which initially launched exclusively for subscribers, is now available to a broader audience at no cost, starting from Thursday.

    Gemini Live serves as Google’s response to OpenAI’s Advanced Voice Mode for ChatGPT. This feature empowers users to interact with the chatbot in real-time using spoken language, rather than typing text inputs.

    To access Gemini Live, simply launch the Gemini app and tap the Sparkle icon located at the bottom right corner of the screen. After your conversation with the AI, you have the option to end the session by pressing the Stop button or by saying, “stop.” Following this, a transcript of your conversation will be generated and saved for future reference in your chat history.

    However, it’s important to note that there are some restrictions. Currently, Gemini Live is available only to users who speak English and use the Android version of the app. It is not compatible with iOS devices or other Workspace applications like YouTube Music and Gmail, though support for these platforms is anticipated in the future.

    In contrast, OpenAI’s Advanced Voice Mode is still in beta testing and has only been made accessible to a limited group of ChatGPT Plus subscribers. While OpenAI plans to extend this feature to all its subscribers, no specific timeline has been provided. Users interested in this feature will need to subscribe for $20 per month without guaranteed access to the rollout.

    Both Google and OpenAI are exploring ways to integrate mobile device cameras into their voice chat functionalities, potentially allowing for enhanced multimodal interaction when responding to spoken queries. However, no exact release dates for these features have been announced by either company.

    Interested in trying out Gemini Live? You can download the Gemini App from the Google Play Store.

  • Blackstone Acquires Australian AirTrunk for $16B in Pursuit of AI Growth

    Blackstone Acquires Australian AirTrunk for $16B in Pursuit of AI Growth

    Blackstone, the American private equity firm, announced on Wednesday its agreement to purchase AirTrunk, a data center operator based in Sydney. This acquisition aligns with the firm’s strategy to invest in AI-related assets within the Asia-Pacific region.

    This deal marks the largest leveraged buyout of the year to date, occurring as private equity dealmaking begins to rebound after a significant increase in financing costs during 2022 and 2023. These rising costs had previously made it challenging to finance substantial acquisitions, as reported by Yahoo.

    Working alongside the Canada Pension Plan Investment Board (CPP Investments), Blackstone is acquiring AirTrunk from Macquarie Asset Management (MAM) and the Public Sector Pension Investment Board (PSP).

    In a statement made on Wednesday, CPP Investments noted that upon completion of the transaction, it will hold a 12 percent stake in AirTrunk.

    As foreign entities are involved in the acquisition, the Australian Foreign Investment Review Board (FIRB) will need to approve the deal. With a valuation of $16.1 billion, this acquisition is considered the largest buyout in Australia for the year and among the most significant in recent history.

    Blackstone President Jon Gray remarked that this acquisition represents the company’s “largest investment in the Asia Pacific region,” highlighting that “AirTrunk is a critical move as Blackstone strives to become the leading digital infrastructure investor globally across various sectors, including data centers, power, and associated services.”

  • What is DARPA’s silent talk project?

    What is DARPA’s silent talk project?

    DARPA’s groundbreaking Silent Talk initiative, decoding brain signals to enable speechless soldier communication on the battlefield.

    In a pioneering leap in military technology, DARPA’s Silent Talk project heralds a transformative era in communication. This ambitious initiative aims to revolutionize soldier-to-soldier communication on the battlefield, bypassing vocalized speech through the decoding of neural signals.

    Deciphering Brain Signals: The Future of Communication

    The Silent Talk project hinges on the idea that words exist as distinct neural signals before they are vocalized. By mapping individual Electroencephalogram (EEG) patterns to specific words, DARPA aims to crack this neural code for seamless transmission to an intended recipient.

    Breaking Down the Goals

    Three pivotal objectives steer the project’s trajectory. First, mapping EEG patterns to individual words; second, determining the universality of these patterns among people; and finally, developing a method to transmit these patterns to another individual.

    Transforming Military Communication

    This disruptive technology holds the potential to revolutionize not only military communications but also offers promise for individuals with disabilities. By sidestepping traditional speech-based communication, Silent Talk could open new avenues for interaction and control.

    Advancing Neurotechnology

    DARPA’s Silent Talk project is part of broader initiatives in brain-computer interfaces and neurotechnology. The significant investment earmarked for this endeavor underscores its strategic importance. The program’s budget, slated at $4 million for the upcoming fiscal year, signifies DARPA’s commitment to pioneering futuristic communication methods.

    Beyond Science Fiction

    While resembling science fiction, Silent Talk is a tangible manifestation of advancements in neuroscience and technology. These innovations have the potential to transcend military applications and cater to broader societal needs, offering novel ways for individuals with disabilities to engage with the world.

    Looking Ahead

    As DARPA forges ahead with the Silent Talk project, the realm of communication teeters on the brink of a groundbreaking transformation. The convergence of neuroscience and technology paints a promising picture of a future where words transcend speech, unlocking a new era of communication possibilities.

    DARPA’s Silent Talk project stands at the vanguard of innovation, poised to redefine communication paradigms both on and off the battlefield.

  • Top 5 Best AI Laptops in [year]

    Top 5 Best AI Laptops in [year]

    For those unfamiliar, Intel recently introduced its latest chipsets known as Intel Core Ultra processors. What sets these processors apart? They boast an NPU, promising a significant enhancement for AI and machine-learning tasks on these innovative “AI laptops.”

    Let’s delve into the lineup of the recently unveiled AI Laptops.

    1. Samsung Galaxy Book 4 Ultra

    Samsung Galaxy Book 4 Ultra
    Photo: SamMobile

    The revamped Samsung Galaxy Book 4 Ultra is now equipped with the latest Intel Core Ultra processor series, offering an NPU, with options available up to Core Ultra 9.

    Here’s a breakdown of its specifications:

    • Offers an Nvidia GeForce RTX 4070 GPU at its highest configuration.
    • Supports up to 64GB of RAM.
    • Provides up to 2TB of storage capacity.
    • Boasts a 16-inch AMOLED touch display with 400-nit brightness, 120Hz refresh rate, and a resolution of 2880 x 1800 pixels.

    The previous model, Samsung Galaxy Book 3 Ultra, impressed me immensely as a content creation powerhouse in my role as a laptop reviewer. While it’s premature to determine if the Galaxy Book 4 Ultra will match its predecessor’s performance, once I complete my review, I’ll be sure to share my insights with you.

    This laptop isn’t currently on the market. Samsung has indicated that it won’t be available until January, and even then, if you’re outside Korea, the release might be further delayed. Korea has priority for availability, while other regions will eventually gain access to the Galaxy Book 4 Ultra in due course.

    As of now, Samsung has kept the pricing details under wraps, with no official announcements regarding the cost of the Galaxy Book 4 Ultra.

    2. Asus Zenbook 14 OLED

    Asus Zenbook 14 OLED
    Photo: Tech Advisor

    The Zenbook 14 OLED, another AI-capable laptop, boasts a 120Hz display with a high resolution of 2880 x 1800 pixels (2.8K), offering an impressive visual experience.

    You do have the option to opt for a variant with a 1080p, 60Hz panel, but given Asus’ claim of “lifelike colors” on the 2.8K panel, it might not be the preferred choice for an enhanced visual experience.

    Here are some additional features:

    • Offers configurations with an Intel Core Ultra 9 processor.
    • Equipped with Intel Arc graphics for enhanced performance.
    • Supports up to 32GB of RAM.
    • Provides storage options up to 1TB SSD capacity.

    In addition, it weighs just 2.6 pounds and boasts a slim profile of only 0.5 inches. This compact build houses all that AI capability. However, it’s not yet available for purchase. The Zenbook 14 OLED is expected to hit the market in early 2024, priced at $1,299.

    3. MSI Prestige 13 AI Evo

    MSI Prestige 13 AI Evo
    Photo: MSI

    The Prestige 13 AI Evo’s remarkable lightness took me by surprise during a recent MSI briefing when I had the chance to lift it.

    Another remarkable aspect of this AI-capable MSI laptop is its resilience despite its compact size. With military-grade durability, the Prestige 13 AI Evo can endure accidental drops, offering robust protection against impacts.

    Here are the specifications:

    • 13-inch display available with an OLED screen at 2880 x 1800 pixels and a refresh rate of 60Hz.
    • Configurable with an Intel Core Ultra 7 processor.
    • Integrated Intel Arc graphics.
    • Supports up to 32GB of RAM.
    • Offers storage options up to 2TB SSD capacity.

    4. Acer Swift Go 14

    Acer Swift Go 14
    Photo: PC World

    If the pricing of the previous laptops made you flinch, the Acer Swift Go 14 might bring some relief—it’s priced at less than $1,000.

    Here are the specifications:

    • 14-inch display available with an OLED screen at 2880 x 1800 pixels and a refresh rate of 90Hz.
    • Configurable with an Intel Core Ultra 7 processor.
    • Integrated Intel Arc graphics.
    • Supports up to 32GB of RAM.
    • Offers storage options up to 1TB SSD capacity.

    This AI-capable laptop is currently available for purchase at Acer.

    5. Lenovo ThinkPad X1 Carbon Gen 12

    Lenovo ThinkPad X1 Carbon Gen 12
    Photo: Hardware Zone

    The ThinkPad X1 Carbon series holds a prominent position in the laptop market, often considered among the finest choices for those seeking a high-quality work laptop. Lenovo has recently taken a step forward by introducing an updated iteration: the Gen 12 refresh.

    Lenovo claims that the Gen 12 ThinkPad X1 Carbon is significantly enhanced with its new AI-capable chip, promising improved performance and capabilities.

    Here are the specifications:

    • Configurable with an Intel Core Ultra 7 vPro CPU.
    • 14-inch display available with a 120Hz OLED touch panel, boasting a resolution of 2800 x 1800 pixels.
    • Supports up to 64GB of DDR5 RAM.
    • Offers storage options up to 2TB SSD capacity.
    • Features Intel Arc graphics.

    Experiencing the ThinkPad X1 Carbon Gen 12 firsthand, I was struck by its diminished size and slimmer build, a result of its reduced dimensions. Once more, this laptop, powered by a new NPU, delivers an enhancement for AI applications like Stable Diffusion.

  • 5 Reasons Why Are AI Content Writing Tools Bad For SEO?

    AI content writing tools have become increasingly popular. Machines are slowly but surely getting better at writing. They can now help write your emails for you. They can write articles and reports for you. But just because machines can, it doesn’t necessarily mean they should.

    (المزيد…)
  • Chanel AI-Powered Lipstick App Lets You Try Any Shade of Cosmetics

    Recently, Chanel has released an app called the Lipscanner (also known as Chanel AI-Powered Lipstick App), which helps the users find the exact shade through Artificial Intelligence.

    (المزيد…)
  • Google AI Will Allow Visually Impaired People To Go For a Run

    Artificial intelligence will allow the visually impaired and the blind to run independently. The system is being tested by Google, to guide the user will be a Google AI app that transforms the smartphone and a wearable device into a sort of eye and ear that suggests the path, together with a harness and headphones Project Guideline, so it is called, is an early-stage research program.

    (المزيد…)
  • You Can Now Search Songs On Google By Humming or Whistling

    Have a song that you don’t know the lyrics for or the name of the song? Now you can search songs on Google by humming and whistling.

    (المزيد…)