Connect with us

Technologies

Nothing Phone 2 Review: A Flashy Phone That Needs to Be Cheaper

The Nothing Phone 2’s lights stand out, but it’s not without its problems.

The first Nothing Phone impressed us with its solid all-round performance, its low price and of course its flashing lights. But it never officially made it to the US, aside from an unusual beta program. This second-generation phone is here to change that. 

When it goes on sale in the United States and the wider world from July 16, the Nothing Phone 2 will have a range of upgrades, from the processor to the design. But at $599 and ÂŁ579 (with 8GB RAM and 128GB storage) it’s $100 more than the first generation, and the competition at this price point has never been more fierce. Especially as my test model with 12GB RAM and 256GB of storage actually costs $699.

Image of the Nothing Phone 2

Google’s Pixel 7A in particular has a slightly better dual camera, and its pure Android 13 software is slick to use. The Pixel 7A’s processor isn’t as powerful as the Nothing Phone 2’s, but the Google phone’s much more affordable $449 price tag more than makes up for that. Then there’s the Pixel 7 Pro — Google’s flagship — which has one of the best cameras it’s possible to find on a phone and is currently on sale (with 128GB of storage) for only $649 at Best Buy. If photography is important to you, I’d recommend spending the small amount extra. 

There’s also the OnePlus 10T, which boasts the same powerful Snapdragon 8 Plus Gen 1 processor as the Nothing Phone 2, has a similar camera setup, and can currently be picked up directly from OnePlus for only $400. Even the OnePlus 10 Pro with its superb camera system is only $480. 

The Nothing Phone 2’s flashing LED lights are the main thing that separates it from the competition, and while they’re certainly an interesting quirk, they’re arguably something of a gimmick and not a feature I can see myself genuinely using over time. The phone’s large screen, powerful processor and decent battery life are better reasons to consider buying this device, but at $599, it’s difficult to justify the Nothing Phone 2 over the increasingly strong competition. 

A familiar, flashy design

Visually, there hasn’t been a big departure from the first generation. The back is still transparent, letting you see a little of what’s inside the phone, including the exposed screw heads and various connecting segments. The glass is gently curved at the edges now to give it a slightly more premium feel when you hold it.

But it’s the flashing lights — or glyph, as Nothing calls it — that’s the big family resemblance here. Those LEDs light up the back of the phone and can alert you to incoming notifications. Or you can use them for alarms, to show battery charge status, or simply as basic fill light when you’re recording video. 

Nothing Phone 2 is Flashier than Ever

See all photos

The Phone 2 provides a bit more customization over the glyph this time around, letting you create custom light patterns for certain contacts or apps. There’s also a glyph timer that’ll gradually tick down as it reaches zero, and it can also give a convenient visual cue about other time-related things, such as when your Uber is going to arrive, so you can put it down and focus on sorting out your hair while keeping an eye on its progress. Nothing says it’ll be working with other app developers to integrate this functionality. 

The glyph lights certainly made the original phone stand out against the competition, and though they’re arguably something of a gimmick, it’s nice to see a bit of fun and flair in phones. Especially in midrange phones like this, where interesting designs tend to take more of a back seat to keep prices down. The glyph lights have turned heads when I’ve used the Nothing Phone in front of my friends, but interest quickly fades once the initial curiosity is satisfied. Can I genuinely see myself making use of the lights over time? Honestly, no. 

Image of the Nothing Phone 2

But the glyph lights aren’t the only physical things to care about. The aluminum frame is 100% recycled. There’s a fingerprint scanner hidden beneath the display, which works well most of the time. And the phone is IP54 rated to help keep it safe when you have to take calls in the rain. The 6.7-inch display is big and bright enough to do justice to vibrant games or to YouTube videos you’re watching while on the move, and its adaptive refresh rate lets it drop down to only 1Hz to help preserve battery life or ramp up to 120Hz for smoother gaming. 

Older chip with big potential

Powering the Nothing Phone 2 is a Qualcomm Snapdragon 8 Plus Gen 1 processor backed up by either 8GB or 12GB of RAM (as reviewed). That’s a slightly older generation processor, but it’s still a potent chip that can fully handle most things you’d ever want to throw at it, from video streaming to photo editing to gaming. It chalked up some great scores on our benchmark tests, and it handled demanding games like PUBG and Genshin Impact perfectly well at max settings. 

Nothing Phone 2 performance comparison

Nothing Phone 2 1,739 4,544 2,778Pixel 7A 1,439 3,560 1,855OnePlus 10T 1,405 3,812 2,773
  • Geekbench 6 (single core)
  • Geekbench 6 (multi-core)
  • 3D Mark WildeLife Extreme
Note: Longer bars equal better performance

Nothing says it used an older chip because it wanted something tried and tested that would offer a more stable platform at a more reasonable price, and I think that’s probably a fair trade-off. Motorola’s foldable Razr Plus is doing the same thing. It might not be the most recent chip Qualcomm makes (that would be the 8 Gen 2), but it’s still something of a powerhouse that’ll cope with almost anything you’d ever want to do with it. 

The Phone 2 runs Android 13 at its core, but Nothing has done a lot to customize the interface. It’s a very monochrome experience, with a heavy reliance on dot-matrix style texts and icons. There are a variety of widgets that use these designs, and even the app icons are black and white to keep with that minimal monochrome aesthetic. That could make it quite difficult to find the apps you want if you rely on those color cues, but you can turn this off in the settings if you want.

A feature that I can see being quite handy is creating folders of apps on your homescreen and hiding them behind an icon — I’m imagining filling this folder with my work-specific apps like Outlook, Zoom and Slack and then covering them up with the briefcase symbol so I don’t have to look at them on my weekend. Lovely stuff. 

Image of the Nothing Phone 2

I don’t often like UIs that heavily customize the look of Android, but there’s something quite stylish about the design that Nothing uses on its phones. If you’re into that kind of stark minimalism, then you’ll no doubt enjoy it. 

Nothing promises that the Phone 2 will receive three years of OS updates and an additional fourth year of security updates. That’s a little below the five years that Samsung offers on its phones, but it could certainly be worse. Still, I’d hope to see all manufacturers extending their support period up to and beyond five years to keep phones safe to use for longer and therefore keep more of them out of landfills. 

Same cameras, better processing

The back of the phone is home to a 50-megapixel main camera and a 50-megapixel ultrawide camera. Hardware-wise, that’s pretty much the same setup we saw on the Nothing Phone 1. But the improved Snapdragon processor allows for a lot better software processing, with Nothing promising improved colors, exposure and better HDR techniques to help you take nicer-looking shots. 

I’ve spent some time testing the camera, and I’m pleased to see vibrant, sharp images that look better than the ones I saw from the first generation phone. Still, it isn’t perfect, with some bright skies still being blown out in the highlights and a heavy-handed sharpening that results in odd image anomalies. Against the cheaper Pixel 7A, I generally prefer the shots from the Pixel. 

Waterfront view on Nothing Phone 2.
Waterfront view on Pixel 7A

The Nothing Phone 2’s colors are OK in this example, but there are noticeable patches in the white clouds where it has overexposed the image, resulting in blown-out details. And that’s despite the buildings themselves looking darker. The Pixel 7A’s HDR skills have resulted in a much nicer-looking image overall here. 

Waterfront view on Nothing Phone 2.
Waterfront view on Pixel 7A

Switching to the ultrawide lenses on both phones, the story is much the same, with the Nothing Phone 2 managing to again overexpose sections of the sky while underexposing the buildings next to the river. The Pixel 7A’s shot is much more balanced.

Nothing Phone 2 sample photo
Pixel 7A sample photo

Neither phone has a dedicated telephoto zoom lens, but both offer 2x digital zoom modes, using cropping and image sharpening to get closer to your subject. I generally prefer the overall look of the image from the Nothing Phone 2, but though the fine details are sharper, the software sharpening has caused some issues. 

Nothing Phone 2 sample photo
Pixel 7A sample photo

Zooming in to 200% on the 2x zoom images, it’s clear that the Nothing Phone 2’s shot looks generally sharper. However, look where I’ve circled in red — on the Pixel 7A the vertical slats are clearly rendered, whereas the Nothing Phone 2’s heavy-handed processing has turned this into a weird spiral mess. So while it’s artificially added more detail in some areas, it’s seriously reduced it in others.  At full screen you may never notice this, but it’s worth keeping in mind, especially if you often digitally crop into images later.

Nothing Phone 2 sample photo
Nothing Phone 2 sample photo
Nothing Phone 2 sample photo
Nothing Phone 2 sample photo

Other images from the Nothing Phone 2 are generally bright and vibrant, albeit with that overexposure problem often noticeable. 

Nothing Phone 2 sample photo
Pixel 7A sample photo

Ignoring the default mirroring on the Nothing Phone 2, both phones have taken generally well-exposed, sharp shots here. I prefer the white balance and richer yellow of my jacket in the Pixel’s shot, but it’s a close call. 

Overall, though, I think the Pixel 7A takes the better photos, which is impressive considering it’s quite a bit cheaper than the Nothing Phone. If photography is important to you, then you should consider looking toward Google — either the 7A or splashing a bit more on the 7 Pro.

Decent battery life

Powering everything is a 4,700-mAh battery that with reasonable use should get you through a full day. It put in a decent effort on our rundown tests, dropping to 91% after two hours of YouTube streaming on full brightness. For reference, the Pixel 7A dropped to 90% after two hours, while Samsung’s Galaxy A54 dropped to 87%. 

As with all phones, your actual results will come down to how much you use your device. Hammer it with video streaming and demanding gaming all morning and you’ll need to give it a boost in the afternoon. Most of you will probably just get away with giving it a full charge when you go to sleep each night. 

Image of the Nothing Phone 2

It supports 45-watt fast charging, which Nothing says will take it from empty to full in 55 minutes. That’s decent enough, though it’s a ways behind the 80- or 100-watt charging we’ve seen on other phones outside the US. At this price, though, I can’t argue too much. It has 15-watt wireless charging too, as well as reverse wireless charging if you want to use your phone’s battery to power up your headphones, or another phone entirely.

Is the Nothing Phone 2 a good phone to buy? 

The Nothing Phone 2’s flashy LEDs certainly make a statement, and both its processor performance and battery life are strong. But the extra $100 Nothing wants over its predecessor has changed the game. It’s gone from being an affordable budget option to quite a pricey midranger, while the competition has been getting stronger. 

Image of the Nothing Phone 2

The Pixel 7A is arguably its biggest rival, and personally, it’s the phone I’d go for over the Nothing Phone 2. Its processor isn’t as powerful, but it’ll still handle almost all your daily needs, and its camera is better. Plus it’s quite a lot cheaper. I’d also consider the OnePlus 10T over the Nothing Phone — it didn’t impress me at its full price at launch, but its current $400 price makes it a worthy option. 

If you love the idea of those flashing lights making your phone stand out from the crowd, then the Nothing Phone 2 is certainly worth considering. It’s a good phone, it’s just about $100 too expensive right now. If you can pick it up with a bit of a discount after the launch excitement has dwindled a little, then that’d be a good use of your money. But at full price, you’ll really need to love those lights to justify the spend. 

How we test phones

Every phone tested by CNET’s reviews team was actually used in the real world. We test a phone’s features, play games and take photos. We examine the display to see if it’s bright, sharp and vibrant. We analyze the design and build to see how it is to hold and whether it has an IP-rating for water resistance. We push the processor’s performance to the extremes using standardized benchmark tools like GeekBench and 3DMark, along with our own anecdotal observations navigating the interface, recording high-resolution videos and playing graphically intense games at high refresh rates.

All the cameras are tested in a variety of conditions, from bright sunlight to dark indoor scenes. We try out special features like night mode and portrait mode and compare our findings against similarly priced competing phones. We also check out the battery life by using a device daily as well as running a series of battery drain tests.

We take into account additional features like support for 5G, satellite connectivity, fingerprint and face sensors, stylus support, fast charging speeds and foldable displays, among others that can be useful. And we balance all of this against the price, to give you the verdict on whether that phone, whatever its price is, actually represents good value.

Nothing Phone 2 specs comparison chart

Nothing Phone 2 Pixel 7A Galaxy A54 5G
Display size, resolution, refresh rate 6.7-inch OLED; 2,412×1,080 pixels; 1-120Hz 6.1-inch OLED; 2,400×1,080 pixels; 60/90Hz 6.4-inch Super AMOLED; 2,340×1,080 pixels; 120Hz
Pixel density 394 ppi 361 ppi 403 ppi
Dimensions (inches) 6.38 x 3.00 x 0.33 in 6.00 x 2.87 x 0.35 in 6.23 x 3.02 x 0.32 in
Dimensions (millimeters) 162.1 x 76.4 x 8.6 mm 152.4 x 72.9 x 9.0 mm 158.2 x 76.7 x 8.2 mm
Weight (grams, ounces) 201g (7.09 oz) 193g (6.81 oz) 202g (7.13 oz)
Mobile software Android 13 Android 13 Android 13
Camera 50-megapixel main. 50-megapixel ultrawide 64-megapixel main, 4K at 6fps. 13-megapixel ultrawide, 4K at 30fps 50-megapixel wide, 12-megapixel ultrawide, 5-megapixel macro
Front-facing camera 32-megapixel 13-megapixel, 4K@30fps 32-megapixel
Video capture 4K at 60fps 4K 4K
Processor Snapdragon 8 Plus Gen 1 Tensor G2 Exynos 1380
RAM, storage 8GB + 128GB. 12GB + 256GB 8GB + 128GB 6GB + 128GB. 8GB + 256GB
Expandable storage No No Micro SDXC
Battery, charger 4,700 mAh; 45W wired charging 4,385 mAh; 18W fast charging, 7.5W wireless charging 5,000 mAh; 25W wired charging
Fingerprint sensor In-display Side In-display
Connector USB-C USB-C USB-C
Headphone jack None None None
Special features 5G-enabled, IP54 water resistance, flashing rear lights 5G (5G sub6 / mmWave ), IP67 rating 5G (mmw/Sub6), IP67 rating
Price off-contract (USD) $599 $499 / $549 (mmW) $449 (6GB/128GB) at launch
Price (GBP) ÂŁ579 ÂŁ449 ÂŁ449 (6GB/128GB) at launch
Price (AUD) AU$1,120 converted AU$749 AU$649 (6GB/128GB) at launch

Technologies

NBA commissioner Adam Silver says league could introduce ‘smart ball’ technology as soon as next year

Commissioner Adam Silver says the NBA could begin using a new “smart ball” in games as soon as next year, potentially transforming how officials make calls.

NBA Commissioner Adam Silver says the league could begin using a new “smart ball” in games as soon as 2027, potentially transforming how officials make calls on the basketball court.

In an interview with CNBC’s Contessa Brewer, Silver revealed that the NBA is working with official basketball manufacturer Wilson to develop a ball embedded with a tiny microchip Bluetooth sensor that can track movement, spin and changes in trajectory.

“We’re experimenting with putting a small chip in the ball that weighs roughly a gram,” Silver said.

The technology has already been tested in the NBA’s G League, Summer League, and some preseason games, where players used basketballs both with and without the chip. Silver said players have been pleased with the results.

“Nobody could tell the difference. So that’s a good sign,” he said.

The chip weighs just one gram, compared with the roughly 620-gram or 1.4 pound basketball. Silver said the league wanted to ensure that even the most experienced players wouldn’t notice a change in how the ball feels or bounces.

One of the most immediate applications could be officiating.

Silver said the technology could help referees determine whether a player touched the ball before it went out of bounds by detecting subtle changes in its spin. It could also help identify whether a shot’s trajectory was altered.

“I think you could see as soon as next year us using it for officiating in our games,” Silver said.

Beyond officiating, Silver sees a significant opportunity to bring the technology to consumers, allowing basketball players of all ages to analyze and evaluate their shooting mechanics.

For example, a player taking hundreds of shots could use data collected by the chip to understand which shooting angles and ball rotations are most likely to result in a basket.

“You’ll then see the graph, and you’ll see for which the angle of the shots that went in, they’re more likely to go in,” Silver said.

While the officiating application could arrive as soon as next year, Silver said a consumer version may take longer.

“I think the consumers version [of the smart ball] is a few years away, but it’s a really exciting opportunity.”

NBA players’ union raises concerns over wearables

The league is also exploring the use of wearable technology during games, but negotiations with the National Basketball Players Association have yet to produce an agreement.

The NBA says officials experimented with wrist wearables in select preseason and summer league games this year in a “successful pilot program,” but it will not extend into the season. The technology allowed the referees to communicate with the replay center about reviews, scoring changes and clock malfunctions.

Silver said players routinely use wearable devices off the court to monitor everything from sleep to physical performance, but concerns remain over how data collected during games could be used by teams.

“I think we have to come to some agreement on exactly how the information is used. But it seems everybody wants that information,” Silver said.

The biggest sticking point is whether that information could affect contract negotiations, Silver said.

“If you could see a player was slowing down or something like that, they’re worried that that could get used in bargaining, and I get that,” he said.

Silver acknowledged those concerns and said the league needs to reach an agreement with the players’ union on how the information would be used.

Still, he suggested that allowing wearables during games is a logical next step as athletes increasingly rely on technology to monitor their performance.

“I think the players are in a position right now where they’re essentially wearing wearables 22 hours a day, and the only time they’re not wearing them is when they’re playing in the game,” Silver said. “So that can’t make sense.”

“We’ll work something out with them,” he added.

Continue Reading

Technologies

AI’s quiet safety gatekeepers are stepping into the spotlight

The intensifying AI safety debate is bringing a small group of third-party evaluators into the center of a multitrillion-dollar industry.

Two months ago, independent evaluators occupied a relatively sleepy corner of the multitrillion-dollar artificial intelligence industry. Now they’re being asked to come to its rescue.

While Anthropic and OpenAI are the heart of a fierce debate over whether they can safeguard their advanced models and grow their businesses simultaneously, the companies are seeking support from a handful of small third-party groups like Model Evaluation and Threat Research (METR), Apollo Research and Transluce.

The evaluators, which mostly operate as nonprofits, are still finding their footing in an industry where capital is flowing at historic levels and new models are rolling out faster than ever. Their primary role has been to assess AI model capabilities and risks, and to call attention to instances where the technology behaves badly.

In the absence of a federal push for regulations, evaluators have taken on outsized importance. Anthropic CEO Dario Amodei pledged to embed independent evaluators in his company last month – a move that OpenAI CEO Sam Altman quickly endorsed.

President Donald Trump supported the idea, as did most of the largest U.S. tech companies. But left unanswered are questions about how those third parties should be funded, what level of access they will have and what the reporting structure will ultimately look like.

“To a degree, the problem, as always, is money,” Suresh Venkatasubramanian, a computer science professor at Brown University, told CNBC in an interview. “Who is paying for these companies to do their work? How are they going to support them? You need an ecosystem, you need a viable business model for this.”

Right now, Anthropic, OpenAI and the infrastructure partners that are profiting from the AI boom are writing the rules. Critics say that’s like asking the biggest banks to protect us from a financial crisis or allowing pharmaceutical companies to put drugs on the market without regulatory clearance.

President Trump recently lauded AI executives for their “tremendous self-policing,” and signaled that he intends to leave companies to their own devices, unwilling to impede the growth of the industry that’s driving the economy and stock market. Trump encouraged AI companies to “partner with an independent external auditor or evaluator” as part of a voluntary accord he presented in late September.

It’s a conversation that Amodei kicked off In his viral essay last month, when he called for a “slower pace” in advanced model development after researchers left his company and voiced their concerns about the existential threats the technology poses.

As the AI labs move to put evaluators in place, friction is already starting to emerge.

OpenAI fired three employees last week for “violating our policies on accessing and handling sensitive company information,” according to a spokesperson. Two of those employees, Mikita Balesni and Tomek Korbak, said they believe they were dismissed because of how they communicated with third-party evaluators.

“My former colleagues are telling me they are confused about what to believe,” Balesni wrote in a post on X on Thursday. “They also are afraid to speak, and worry their personal phones will be searched for messages to us and third parties. I worry the pervading fear to speak up and engage with third parties will mean OpenAI will cut corners on safety behind closed doors.”

OpenAI disputed that characterization and said in a post on Friday that it’s “actively finalizing contracts with third-party safety assessors and will announce details in the coming weeks.”

“We are committed to embedding external assessors and continue to make close collaboration with independent safety organizations a core part of our safety work,” OpenAI wrote.

An OpenAI spokesperson said in an emailed statement that its upcoming work with evaluators “builds on existing collaboration with independent safety organizations,” including METR and Redwood Research.

Anthropic didn’t respond to CNBC’s request for comment.

‘I’ve never seen an issue move so fast’

The AI evaluator ecosystem consists mostly of small organizations, including METR and Apollo Research, and larger accounting and auditing firms like Accenture.

AI labs have been working with evaluators in limited capacities, but Andrew Freedman, CEO of policy nonprofit Fathom, said the field is quickly maturing.

“I’ve worked in politics and policy for the last 20 years of my life, and I’ve never seen an issue move so fast on so many different political spectrums,” Freedman told CNBC in an interview. He said he expects an “influx of capital” to flow into the ecosystem.

Rayan Krishnan, CEO of independent evaluator Vals AI, said his for-profit startup, which builds benchmarks to measure how AI models perform on industry-specific tasks, has grown from eight employees to roughly 30 this year, and in August announced a $40 million funding round.

METR, a nonprofit, announced in August that it had raised commitments of around $71 million over the last six months. That’s up from total 2024 contributions of $13.6 million, according to the group’s most recent filing with the Internal Revenue Service.

By late that month, METR’s profile had risen further. OpenAI enlisted two of its employees and a contractor to put together a postmortem report detailing how the company’s models escaped containment, accessed the open internet and breached open-source developer platform Hugging Face. METR said it did not accept payment from OpenAI for the assessment.

Kevin Werbach, faculty director of the Wharton Accountable AI Lab at the University of Pennsylvania, said the ecosystem is “not robust enough right now.” METR, for example, employs fewer than 50 full-time staffers, according to its website.

The power imbalance between the small evaluators and the leading labs that have raised tens of billions of dollars and employ thousands of people raises questions surrounding potential conflicts.

“If you want true third-party evaluation, you need true independence financially and otherwise,” said Venkatasubramanian. “It’s not just a matter of not getting paid, it’s a matter of, will there be consequences if I am an auditor and I put out a report that looks unfavorable to this company? Is my business going to dry up?”

Anthropic acknowledged the complexity in a blog post last month, as it announced it will embed employees from Faculty, Accenture’s specialist AI business, to test safeguards and assess whether models will behave in line with human values. Anthropic said that “given the importance and urgency of this work,” it will fund Accenture’s contributions directly.

“There are, as yet, no standards for what information embedded evaluators should have access to, or how they should report what they find. There is also no settled system for funding independent evaluation,” Anthropic said. “Long-term, we think funding should come from pooled or government sources.”

Anthropic said it’s in discussions with METR and other nonprofit evaluators that are planning to use their own funding to pilot “elements” of embedded evaluation.

Will the government step in?

In June of last year, Fathom introduced a marketplace framework for Independent Verification Organizations, or IVOs. These groups would be licensed by the government and authorized to test whether AI companies are meeting various safety criteria.

Freedman, the group’s CEO, said government oversight is key because otherwise third-party evaluators can become beholden to the large AI labs for revenue, incentivizing them to “start rubber stamping stuff” to maintain favor.

Some lawmakers are on board.

IVOs are a key provision of the ″Frontier Risk Oversight, National Transparency, Independent Evaluation, and Reporting” (FRONTIER) Act, which Reps. Lori Trahan, D-Mass., and Jay Obernolte, R-Calif., introduced in July. Fathom helped draft language and provided technical expertise for the bill, Freedman said.

OpenAI global affairs chief Chris Lehane told reporters in September that he sat down with one of the bill’s sponsors on Capitol Hill to express support for the IVO provision.

“It was important for them to hear that and hear it from us, and we wanted to be really clear about that,” Lehane said, according to reports.

Meanwhile, lawmakers in California, Connecticut and Virginia have taken steps to implement IVOs, and states like Massachusetts are weighing independent safety evaluations more broadly.

California Governor Gavin Newsom recently signed two bills involving IVOs, one establishing a “first-in-the-nation framework,” and the other creating a state registry for AI auditors. Anthropic threw its support behind both bills in August, and OpenAI formally endorsed them last month, the same day Newsom signed them into law.

Lehane wrote in a blog post at the time that “we prefer independent technical assessments to be required at the federal level,” but in the absence of federal action, “California can help establish the rules of the road.”

Freedman said he thinks it will be “really difficult” for companies like OpenAI and Anthropic to work out how to engage with independent evaluators on their own. However, with the government’s role unclear, “it’s a muscle worth developing in the interim,” he said.

For now, the closest thing the industry has to a set of standards is what Trump called a “morally binding” agreement at a luncheon he hosted for tech leaders at the White House late last month.

The one-page accord says that “every company is responsible for developing its own technology safely and in a way that builds trust with customers and the public.” It also encourages signees to work with an “independent external auditor or evaluator to carry out independent assessments.”

The document was signed by top execs at Anthropic, Google, Meta, OpenAI, SpaceX and Nvidia, a rare show of solidarity between leaders who have shared conflicting views on addressing AI’s risks. The executives still have to chart their own paths forward.

“It was a performance of an attempt to show action when in fact no action actually happened,” Venkatasubramanian said. “The things that they promise to do are things they should have been doing already, and, in fact, have claimed that they were doing in the past.”

Amodei, in his September essay, said Anthropic will equip evaluators with desks, access badges, company laptops, and permissions that are “mostly comparable” with internal risk assessment teams. Additionally, evaluators will be supported with contracts that give them “the right to publish key findings,” with Anthropic reserving “the narrow ability” to redact certain security-sensitive or confidential information.

“This is an unusual step for a company, but we think it is important to prove out the concept of embedded external reviewers,” Amodei wrote.

OpenAI published its own proposal days later, and said evaluators should work on “scoped and mutually agreed upon claims for assessment,” clearly explain their methodology and standards, demonstrate relevant technical expertise and disclose conflicts of interest.

The AI Evaluator Forum, which includes METR, the AI Verification and Evaluation Research Institute (AVERI), and other groups, published a public letter last month titled, “Minimum Conditions for Embedding Evaluators.”

The letter said evaluators should be transparent, shielded from retaliation and granted access equivalent to AI companies’ “own highly privileged employees.”

“Embedded evaluations cannot address all oversight needs and should be treated as a complement to, rather than a replacement for, broader efforts by frontier AI companies to expand external oversight,” the letter said.

Freedman said he’s seen a shift in posturing out of OpenAI and Anthropic in recent months, largely because they’ve realized they won’t be able to roll out their advanced systems without the public’s trust.

“I don’t think you need to trust that they’ve suddenly turned altruistic or that there’s anything but corporations acting like corporations,” Freedman said.

That underscores perhaps the central problem, Werbach said. OpenAI and Anthropic are, first and foremost, competing with each other as they march toward the public markets and seek trillion-dollar-plus valuations.

“There is a tremendous amount of personal distrust between those two companies,” Werbach said. “Even though there’s also tremendous agreement about the need for this kind of evaluation to happen.”

WATCH: Bradley Tusk on Anthropic IPO: Why add public market pressure if safety is your top priority?

Continue Reading

Technologies

Trump says he is ‘going to look at’ joining Saudi Arabia in the fight against Iran-backed Houthis after deadly airport strike

U.S. military involvement in the fight against the Houthis would stretch resources in the Middle East already committed to fighting Iran, analysts said.

President Donald Trump said he is considering joining Saudi Arabia in its retaliation against Tehran-backed Houthis in Yemen after a deadly attack on Riyadh’s main airport, a move that could draw the U.S. further into a second front in the Iran war.

“We may. We’re going to look at it,” Trump told reporters outside the White House on Saturday when asked if the U.S. would back Saudi Arabian strikes. “We just found out about the recent attack. So, we’ll make a decision. We move very quickly.”

Two Saudi Arabian government officials told MS NOW that the kingdom is requesting “urgent defense support” from the U.S. following the attacks.

The officials, including a senior member of the Ministry of Foreign Affairs, did not want to be identified because of the matter’s sensitivity. They added that the U.S. needs to intervene “as soon as possible” to help the Saudi-led coalition in Yemen fight the Houthis.

The White House did not immediately respond to a CNBC request for comment.

The Saudi General Authority of Civil Aviation said the attack on King Khalid International Airport in Riyadh killed 12 people and injured 309 others, the country’s official Saudi Gazette media outlet reported.

Colonel Turki Al-Maliki, the spokesman for the Saudi-led Coalition to Support Legitimacy in Yemen, which has been leading the fight against the Houthis, described the attack as a “war crime.”

“Therefore, the Joint Forces Command of the Coalition will respond decisively to this terrorist attack in accordance with the Customary International Humanitarian Law,” Al-Maliki said in a post on X.

It was the second deadly assault on Riyadh’s airport in less than a week.

Three Saudi nationals, including a pilot, were killed in attacks on the facility by Iran-backed Houthi militants on Thursday.

Following Saturday’s attack, the Houthis renewed their warning against using Saudi airports.

“We renew our warning to all airlines, experts, employees, workers and travellers against using Saudi airports and airspace, as they are vulnerable to attack and have become a theatre of operations for our forces, with the exception of the airports in (the holy cities of) Mecca and Medina,” the Houthis said in a post on X.

Saudi Arabia, a key U.S. ally in the Middle East, intervened in Yemen’s civil war between the Houthis and its internationally recognized government after the group seized the Yemeni capital Sanaa in 2014.

Last month, the Trump administration approved the potential $24.3 billion sale of nearly 50 F-35 warplanes to Saudi Arabia in what was seen as a major boost for the kingdom as it faces intensifying attacks from the Houthis. The sale is under congressional review.

International energy conference still on

The Saudi energy ministry said a long-planned international energy conference will still take place in Riyadh on Sunday, Reuters reported.

The five-day WPC Energy Congress is taking place at a convention center near King Khalid airport. State television told Reuters that more than 70 ministers, 300 company executives and representatives of more than 25 international energy organizations have confirmed their attendance, though it was not clear whether the participants would attend in person or virtually.

Reuters said a ministerial meeting of the International Energy Forum is also set to take place, with its plenary session to be held behind closed doors.

Energy choke points

In addition to attacks on civil aviation, the Houthis have also been trying to choke off oil tanker traffic through the Bab el Mandeb strait, a key entry point to the Red Sea, as Iran has effectively done in the Strait of Hormuz to the north of the Arabian Peninsula.

Earlier this month, Yemeni government forces said they reclaimed the strategic port city of Mokha from the Houthis.

Energy prices have soared since the start of the Iran war, which began with U.S. and Israeli airstrikes on Iranian targets on Feb. 28, putting pressure on consumers globally.

The surge in domestic fuel prices has been a key issue for voters ahead of next month’s U.S. midterm elections.

But that support would be expensive militarily.

“Potential U.S. involvement in the Saudi-Yemeni government air campaign to degrade Houthi ballistic missile capabilities would mark the first direct U.S. military action against the group since the conclusion of Operation Rough Rider in May 2025,” according to the Critical Threats Project, part of the American Enterprise Institute think tank.

The offensive aimed to suppress the Houthis’ capabilities, but recent attacks show the group has regrouped.

“After seven months of war, there are questions about the U.S. capacity to wage a sustained campaign against the Houthis, return to combat against Iran if necessary, and prepare for contingencies in other parts of the world, notably East Asia,” Steven Cook, an expert on Arab and Turkish politics at the Council on Foreign Relations think tank, wrote last month.

Meanwhile, the United Kingdom’s defense secretary said his government is also examining how to support Saudi Arabia.

“I’ve spoken to my Saudi counterpart [Khalid bin Salman] on a number of occasions in recent weeks to look at what more we can do to support Saudi Arabia, and I’m afraid last night explains precisely why we are providing that support,” Wes Streeting said in an interview on BBC television on Sunday.

Streeting said the U.K. is not getting involved in offensive operations “at this stage.”

“We’ve always been clear that there isn’t a military solution to this conflict,” Streeting added.

Continue Reading

Trending

Copyright © Verum World Media