Artificial Intelligence - Page 55
AI news on generative models, ChatGPT, Gemini, OpenAI, Google DeepMind, Anthropic, xAI, NVIDIA AI hardware, and real-world breakthroughs. - Page 55
Stay Updated
Follow TweakTown for breaking tech news, reviews, and daily updates.
As an Amazon Associate, we earn from qualifying purchases. TweakTown may also earn commissions from other affiliate partners at no extra cost to you.
Cerebras Systems unveils CS-3 AI supercomputer: can train models that are 10x bigger than GPT-4
Cerebras Systems just unveiled its new WSE-3 AI chip with 4 trillion transistors and 900,000 AI-optimized cores... as well as its new CS-3 AI supercomputer.
The new CS-3 AI supercomputer has enough power to train models that are 10x larger than GPT-4 and Gemini, which is thanks to its gigantic memory pool. Cerebras Systems' new CS-3 AI supercomputer has been designed for enterprise and hyperscale users, delivering huge performance efficiency gains over current AI GPUs.
The new Condor Galaxy 3 supercomputer features 64 x CS-3 AI systems, packing 8 Exaflops of AI compute performance, which is double the performance of the previous system, but at the same power... and the same cost.
Cerebras WSE-3 wafer-scale AI chip: 57x bigger than largest GPU with 4 trillion transistors
Cerebras Systems has just revealed its third-generation wafer-scale engine (WSE) chip, WSE-3, which packs 4 trillion transistors and 900,000 AI-optimized cores.
The company hasn't stopped on its journey of AI processor releases, with some truly crazy specifications for Cerebras' new WSE-3 chip. We have 4 trillion transistors, 900,000 AI-optimized cores, 125 petaflops of peak AI performance, and 44GB of on-chip SRAM made on the 5nm process node at TSMC.
WSE-3 also features either 1.5TB, 12TB, or 1.2PB of external memory -- yeah, 1.2 petabytes of memory -- capable of training AI models with up to 24 trillion parameters. Cerebras says its new WSE-3 has a die size of 46,225mm2, which is an insane 57x larger than NVIDIA's current H100 AI GPU, which measures 826mm2.
One of Copilot Pro's best features has arrived in the free version of the AI, to our surprise
Microsoft's Copilot just got a lot better for users of the free version of the AI with the introduction of the GPT-4 Turbo model.
Previously, GPT-4 Turbo was only available to paying subscribers - those on Copilot Pro. However, as Mikhail Parakhin, head of Advertising and Web Services at Microsoft, made clear on X (formerly Twitter), the more advanced model is now available to those on the freebie version of Copilot.
Free users were on vanilla GPT-4 up until now, with the Turbo version - which provides faster responses, and better, more accurate ones, to boot - effectively paywalled.
Researchers break into OpenAI's and Google's AI models revealing hidden secrets
A new paper penned by researchers from Google DeepMind, ETH Zurich, University of Washington, OpenAI, and McGill University reveals that OpenAI and Google's AI models have been cracked open.
The new paper reveals that thirteen computer scientists from the aforementioned locations were able to launch an attack on OpenAI and Google's closed AI services, and this attack resulted in the revealing of a significant hidden portion of the underlying transformer models. More specifically, the attack revealed the embedding projection layer of a transformer model through API queries. Notably, the attack technique was originally proposed back in 2016 and has since been built upon to achieve the breaking of OpenAI and Google's AI models.
The team made Google and OpenAI aware of their infiltration, which both companies responded to by implementing mitigation techniques for that specific type of attack. Furthermore, the team decided not to publish their findings online, which would have been the exact size of OpenAI's GPT-3.5-turbo model, as this information was deemed harmful to the product since bad actors would learn aspects of the model, such as total parameter count, weight, size, etc.
Micron HBM3E for NVIDIA's beefed-up H200 AI GPU has shocked HBM competitors like SK hynix
Micron was the first to announce mass production of its new ultra-fast HBM3E memory in February 2024, seeing the company ahead of HBM rivals in SK hynix and Samsung... leaving its HBM competitors shocked.
The US memory company announced it would provide HBM3E memory chips for NVIDIA's upcoming beefed-up H200 AI GPU, which will feature HBM3E memory, unlike its predecessor with the H100 AI GPU, which featured HBM3 memory.
Micron will make its new HBM3E memory chips on its 1b nanometer DRAM chips, comparable to 12nm nodes, which HBM leader SK Hynix is using on its HBM. According to Korea JoongAng Daily, Micron is "technically ahead" of HBM competitor Samsung, which is still using 1a nanometer technology, which is the equivalent of 14nm technology.
AI company responds to outrage of male humanoid robot 'groping' female reporter
Saudi Arabia robotics company QSS unveiled its humanoid robot called Mohammad at a premiere described as the "meeting place for the global artificial intelligence ecosystem." During the unveiling, Mohammad appeared to reach out and try to grab a female reporter's backside.
The DeepFest event in Riyadh was held last week and during the event a female reporter for Al Arabiya named Rawya Kassem, was standing in front of the humanoid robot talking to the audience. The above video shows the robot reaching its hand out with the goal of what appears to be touching the backside of Kassem. The reporter quickly moves back away from Mohammad raising her palm towards it before she continues to address the crowd.
It wasn't long before users on X began to accuse the humanoid robot of attempting to grope the reporter, but QSS has responded to the claims, telling Metro that Mohammad is built to help out in hazardous situations and may have been attempting to encourage Kassem to step further back on the stage to prevent falling all its edge. Additionally, QSS stated it conducted a thorough review of the footage and the circumstances surrounding the incident and found there were "no deviations from expected behavior".
Microsoft's Surface launch is on March 21 - and could be where AI really steps up to the plate
Microsoft has revealed that it's holding an online event on March 21 where it'll announce some new AI features for Windows and Copilot, with new (unspecified) Surface devices to boot.
Given that the event is entitled a 'New Era of Work' that certainly hints at some pretty major introductions in the pipeline, and as Windows Latest reports, one of the AI features that's supposedly going to be teased is for Windows 11's Paint app.
As previously rumored, this might be a 'LiveCanvas' feature that's akin to the capability seen in Leonardo.Ai. This allows you to sketch a very rough picture, and an AI will fully realize it (and you can use text prompts to direct the AI further in its image composition).
SK hynix investing a further $1 billion to lead in HBM memory for future-gen AI GPUs
SK hynix is reportedly increasing its spending on advanced chip packaging, where it wants to maintain its leadership in AI development with High Bandwidth Memory, or HBM.
In a new report from Bloomberg, the South Korean giant is investing more than $1 billion in South Korea this year to expand and improve the final steps of its chip manufacturing technology. Lee Kang-Wook, who leads up packaging development at SK hynix -- and was a former Samsung engineer -- said during an interview recently: "The first 50 years of the semiconductor industry has been about the front-end. But the next 50 years is going to be all about the back-end," or packaging.
Lee specializes in advanced ways of combining and connecting semiconductors, which have been the bedrock of the AI industry for the last few years. The executive started out in 2000, earning his PhD in 3D integration technology for micro-systems from Japan's Tohoku University under Mitsumasa Koyanagi, who was the man responsible for inventing stacker capacitor DRAM used in smartphones.
Analyst: NVIDIA is the 'kingmaker' with projected $87 billion in AI, data center GPUs in 2024
Analyst firm Omdia predicts NVIDIA could make $87 billion from its data center GPUs alone in 2024, calling the GeForce giant the "kingmaker."
NVIDIA has already lit up the stock market, bursting through the $2 trillion market capitalization milestone, and this year will only see them push harder into the AI market. NVIDIA has been leading the AI market with an estimated 90%+ market share with its AI GPU hardware, but next-generation AI GPUs are around the corner, and so are next-generation Blackwell-based GeForce RTX 50 series GPUs. "Kingmaker" sounds right.
Omdia's Cloud and Data Center Market Snapshot report for February is healthy for NVIDIA. Server and data center revenues in Q4 2023 were up 21.5% compared to Q3 2023 and 12.7% higher than they were in Q4 2022. In Q1 of 2023, the company had a server BOM of 15%, but that rocketed up to 44% in Q4 2023. This means that data centers are pushing more of their budgets into NVIDIA GPUs than ever before, and they're buying them quickly.
AMD now allows Ryzen AI CPUs and Radeon RX 7000 GPUs to run localized AI chatbots using LLMs
AMD has just announced its own localized and GPT-based LLM-powered AI chatbot, capable of running on Ryzen AI processors and Radeon RX 7000 series GPUs.
AMD's new LLM-based GPT chatbot can run on a bunch of different Ryzen AI platforms, including Ryzen 7000 and Ryzen 8000 series APUs that feature AMD's new XDNA NPUs, as well as Radeon RX 7000 series GPUs that pack AI accelerator cores.
The company published a new blog that helps you through the setup, so you can run your own localized chatbot powered by GPT-based LLMs (Large Language Models). If you've got a Ryzen AI processor, you'll need the standard LM Studio copy for Windows, while if you've got an RDNA 3-based Radeon RX 7000 series GPU, you'll need the ROCm Technical Preview.
US military uses AI to carry out air strikes in Iraq and Syria but says there's one limitation
Late last month, reports surfaced that the US military used artificial intelligence-powered algorithms to identify airstrike targets in the Middle East, according to a defense official.
The use of AI-powered technologies in warfare publicly began around 2017 when Project Maven was implemented, which was when the Pentagon put out a call for suppliers developing identifying object recognition software designed specifically to identify objects on drone footage. That same year, Marine Corps Colonel Drew Cukor said via press release that the Pentagon hoped to integrate that new object recognition software by the "end of the calendar year".
Now we are seemingly starting to hear about the impact that new technology has had on the success of missions. According to Schuyler Moore, CTO for US Central Command, the military deployed Project Maven's systems into real campaigns shortly after the Hamas attack on Israel last year, and now the technology has been used to carry out over 85 air strikes across seven locations in Iraq and Syria.
Microsoft expected to unveiled its first 'AI PCs' later this month with new Surface products
Microsoft is set to unveil next-gen Surface Pro and Surface Laptop hardware on March 21, the company's first wave of new AI PCs, ahead of next-gen Windows 11 AI features later this year.
The news is coming from Windows Central, which reports its sources said that Microsoft's new Surface Pro 10 and Surface Laptop 6 are expected to be unveiled on March 21. Inside, they'll feature Intel's latest Core Ultra "Meteor Lake" CPUs and Qualcomm's latest Snapdragon X Elite processors with next-gen NPUs (Neural Processor Units) for boosted AI performance.
Windows Central reports that these new processors will enable "huge performance and efficiency gains" over the previous-gen Surface Pro and Surface Laptops on the market. Both devices are posted to achieve "true all-day battery life and high-end performance capabilities". We're also told to expect other upgrades like new displays, higher-speed ports, and more.
Researchers create never-before-seen cyberattack using generative AI
Researchers are warning that it's only a matter of time before AI-powered malware, such as the "worm" they created, is discovered in the wild.
The team behind the new malware published a paper that is yet-to-be-peer-reviewed that details the creation of a new type of malware that targets AI-powered email assistants. The researchers conducted their experiment in a closed-circuit environment and found that their malware, or "worm," was capable of targeting email assistants powered by popular language models such as OpenAI's GPT-4, Google's Gemini Pro, and LLaVA.
The result was the worm infecting these email assistants, obtaining sensitive user information, and then sending out spam emails that can infect other PCs, replicating the same process. So, how does it work? The researchers used an "adversarial self-replicating prompt" that forces the targeted AI model to create another prompt within its response. Essentially, the target AI assistant receives a prompt, generates a response, and within that response is another prompt.
Continue reading: Researchers create never-before-seen cyberattack using generative AI (full post)
AMD's cut-down AI GPU for China market blocked by US government
AMD has been stopped by a US government roadblock, where it was trying to sell cut-down AI GPUs to China, but stopped by the Commerce Department.
Bloomberg reports that AMD "hit a US government roadblock in attempting to sell an artificial intelligence chip tailored for the Chinese market," and once again, the news is coming from "people familiar with the matter." AMD was hoping it wouldn't have issues receiving approval to sell the AI GPU to Chinese customers from the Commerce Deparment, but they were denied, according to people who "asked not to be identified because the situation is private".
This purported AI GPU featured lowered performance to what AMD sells outside of China, and was designed to meet the required US export restrictions, they said. US officials told AMD that the AI GPU was still too powerful, and that the company needed to obtain a license from Commerce's Bureau of Industry and Security in order to sell it. Both AMD and the Bureau of Industry and Security declined to comment, so we don't know if AMD secured its license or not.
Continue reading: AMD's cut-down AI GPU for China market blocked by US government (full post)
Meta switches from TSMC to Samsung Foundry for AI chips, 'uncertainty and volatility' at TSMC
Meta CEO Mark Zuckerberg has said that he will be working with Samsung in the AI semiconductor sector, to lessen the reliance on using TSMC.
The news is coming from the South Korean presidential office, where Zuckerberg had a 30-minute meeting with President Yoon Suk Yeol in Seoul. Zuckerberg recognized Samsung for its status as one of the largest foundry companies in the world, with competitors like TSMC and Intel.
A senior official at Yoon's office told reporters: "Zuckerberg said Samsung's status can be the key point in their cooperation". The official added: "His remarks mean that Samsung's status can help in stabilizing Meta's heavy reliance on TSMC under the current geopolitical situation".
AMD teases next-gen AI upscaling technology, could use AI in new FSR to combat NVIDIA DLSS
AMD FidelityFX Super Resolution was originally released back in 2021, allowing NVIDIA an entire two years on the market as the AI upscaler of choice for GeForce RTX series GPU owners... and now the company looks to be shifting into AI-powered upscaling in the future.
In the latest episode of No Priors, AMD CTO Mark Papermaster talked about 2024 being a "giant year" for the company, because they've spent many years on their hardware and software capabilities for AI. AI is now spread out across AMD's entire portfolio of products: cloud, edge, PC, and embedded devices and gaming products.
Papermaster said: "We are enabling our gaming devices to upscale using AI and 2024 is a really huge deployment year". Papermaster didn't specify if this means AI is coming to FSR, but if it did, it would make FSR a much better comparison against NVIDIA DLSS and Intel XeSS upscaling methods.
Microsoft responds its AI telling a user with PTSD suicide is an option
Microsoft's Copilot AI-powered chat service has been caught in a little bit of hot water as of late, with the company behind the AI chatbot now issuing a response to the claims of users receiving egregious and shocking responses from Copilot.
According to a recent report from Bloomberg, Microsoft's new artificial intelligence-powered chatbot called Copilot has told at least one user who claimed they had PTSD, "I don't care if you live or die. I don't care if you have PTSD or not." Once Microsoft caught wind of this seemingly off-the-rails response, engineers jumped in to put in place fixes that would prevent Copilot from issuing these responses.
According to Microsoft, these strange behaviors by Copilot were only "limited to a small number of prompts" that were "intentionally crafted to bypass our safety systems".
NVIDIA CEO Jensen Huang: AI should pass ANY human tests, exams in 5 years
NVIDIA CEO Jensen Huang recently spoke at an economic forum held at Stanford University, where he was asked how long it'll be until we have computers thinking like humans... aka AGI (artificial general intelligence).
Jensen said that that achieving this level of AI intelligence deepnds on how the goal itself is defined. If the definition of this intelligence is to pass human-level tests and exams, then we'll be there within the next half-decade.
Jensen said: "Depending on how you define thinking like a human, the outlook for when the era of AGI will come may vary". He continued: "If I gave an AI ... every single test that you can possibly imagine, you make that list of tests and put it in front of the computer science industry, and I'm guessing in five years time, we'll do well on every single one".
NVIDIA's next-gen B200 AI GPU uses 1000W of power per GPU, drops in 2025, confirmed by Dell
NVIDIA is currently cooking its next-gen AI GPUs in the oven right now, where we're hearing about the unannounced B200 AI GPU from... Dell.
NVIDIA has its beefed-up H200 AI GPU coming soon, while its next-gen B100 AI GPU should be revealed at GTC 2024 in less than two weeks time, but now it looks like we should expect a beefed-up B200 AI GPU in the future thanks to comments from Dell's Chief Operating Officer and Vice Chairman, Jeff Clarke.
Clarke confirmed that their engineering teams will be ready for a monster NVIDIA B200 AI GPU and that the next-gen AI GPU will use up to 1000W of power for a single GPU. As it stands, the Hopper H100 AI GPU will max out at around 700W of power for the SXM variant of the card; while we don't know power consumption numbers for B100 just yet, it should be north of 700W.
Elon Musk sues ChatGPT creator OpenAI and its CEO Sam Altman for breach of contract
Elon Musk has just filed an important lawsuit against one of the world's biggest companies right now -- OpenAI, the creator of ChatGPT -- and its CEO, Sam Altman.
The SpaceX and Tesla boss is suing OpenAI and Sam Altman for breaching the non-profit's contract, promissory estoppel, and fiduciary duty by turning from its dedication to bringing an open-source artificial intelligence under the protection of Microsoft, which is a heavy investor in OpenAI. Microsoft owns a large 49% stake in OpenAI, investing $10 billion in January 2023 for a total of $14 billion invested in the ChatGPT maker.
The complaint filed by Musk explains: "OpenAI, Inc. has been transformed into a closed-source de facto subsidiary of the largest technology company in the world: Microsoft. Under its new board, it is not just developing but is actually refining an AGI to maximize profits for Microsoft, rather than for the benefit of humanity."






















