Technologies
The Galaxy Watch 8 Pissed Me Off, but I’d Still Recommend It
Samsung’s Running Coach questioned my running skills. But Gemini may have just restored my faith in voice assistants.
The Running Coach on the Galaxy Watch 8 needs to be kicked to the curb. I’m not expecting an Olympic endorsement deal anytime soon, but after 20 years of running (four half marathons, multiple 10K and 5Ks), I’d hope to graduate beyond “beginner.” Not according to Samsung’s latest watch. Either it’s using a rigid set of criteria to assign training plans, or it’s gaslighting me on purpose to tap into my competitive streak. Whatever the case, Running Coach left me questioning its usefulness and cast a gray cloud over my running experience. Something seemed off, so I checked in with Samsung and am still waiting to hear back.Â
Running Coach aside, the $350 Galaxy Watch 8 ($50 more than last year’s Galaxy Watch 7) gets a lot of other things right, and I still recommend it to anyone looking for a solid Wear OS smartwatch. One of the biggest surprises: Gemini. This is the first smartwatch to come with Google’s AI assistant built in, and the voice assistant actually feels useful on the wrist. It’s also one of the most comfortable watches I’ve ever worn (though not the most stylish). It has nearly every feature I could hope for, including a screen that’s blindingly bright and new health sensors for more accurate health tracking.Â
Pros
- Dual sizing options that fit well on smaller wrists
- Comfortable, lightweight design
- Gemini assistant is fast and genuinely helpful
- New health sensors offer more accurate insights
- Bright display is visible in direct sunlight
Cons
- Price is $50 more than the Galaxy Watch 7
- Squared frame isn’t for everyone
- Health features require manual setup
- Running Coach accuracy is questionable so far
- Proprietary straps limit options from third parties
- 40mm model tops out at 30 hours battery life
From feature-rich smart rings (Samsung’s Galaxy Ring included) to budget smartwatches like the $80 Amazfit Bip 6, the competition for your health data is getting fierce. In a crowded landscape, Samsung positions the Galaxy Watch 8 as a high-end alternative with the goal of long-term success: slowing the hands of time, promoting healthy aging and delivering more meaningful measurements.
The result is a mature smartwatch that goes above and beyond the basics, offering new metrics for cardiovascular health, a skin-based antioxidant index, improved bedtime guidance, and yes, a personal running coach that promises to get you “marathon-ready.” I swear I’m not bitter. Most of these tools rely on Samsung’s advanced BioActive sensor, which is available only on the Series 8 models (and the Ultra), and one of the main reasons why you’d consider upgrading. It’s also worth noting that none of these features are medical-grade devices and therefore should be taken with a healthy grain of salt.
After wearing the Galaxy Watch 8 for less than a week, some of the new features still feel like works in progress while others show real potential. Paired with a Galaxy phone, the Watch 8 feels like a confident, integrated health and fitness companion with a voice assistant that might actually talk you into keeping it on.
The Galaxy Watch 8 is available now for preorder for a base price of $350 for the 40mm model, and $380 for the 44mm version. Add $50 more for LTE on either size.Â
Galaxy Watch 8 Running Coach
As a longtime runner, I was genuinely excited about the new Running Coach â a virtual coach that would give me personalized training plans and real-time feedback to whip me back into racing shape. The setup involved filling out a brief questionnaire on my phone about my running and workout habits. Then it asked me to record my longest run in the last three months, which happened to be a 5K.
I’m a no-frills runner; I usually have about 30 minutes to squeeze in a jog, which means getting out the door without searching for a headset or curating the perfect playlist. So the idea of needing headphones just to hear the Running Coach felt like a drag. A quick “turn up the volume to max” command to Gemini saved the day. Fortunately for me, the watch plays the prompts through its speaker, which, while not particularly loud, was loud enough for me to finish the assessment without headphones.Â
The test started with a short warmup, then moved into intervals: a normal pace, an all-out sprint, then back to normal, followed by a cooldown to gauge how quickly my heart rate recovered. In total, it took about 14 minutes. The voice was definitely robotic â not exactly the tough-love human sounding coach I had imagined.
I was still recovering from intense travel and a lingering ACL injury, so I wasn’t expecting a gold star. But with an average pace of 9:45 per mile, I figured I’d at least score higher than level one. Being labeled a beginner and assigned a plan to “build up to a 5K” felt borderline insulting, especially considering I’d just told it that I’d already completed one.
Looking closer at the plan, I saw it had me walking for 30 minutes during the first week, with a goal of running 0.93 miles in less than 10.5 minutes by week four. Both of which I’d already done during the initial assessment.
Meanwhile, a colleague who isn’t a runner and walked the entire test got the same training plan I did. That raised some serious questions. How “personalized” can this really be if two people with vastly different running backgrounds are given the exact same plan?
For now, the experience has left me skeptical â and has definitely taken some shine off a feature I was really hoping to love. It’s possible the coach will recalibrate my training plan as it gathers more running data, but it’s also just as likely that Running Coach itself needs to step up its game with future updates.
Galaxy Watch 8 Antioxidant Index
Samsung’s new Antioxidant Index, which measures carotenoid levels in the skin, is arguably one of the most interesting features on the Galaxy Watch 8, and one of the most confusing.
I didn’t know much about antioxidants beyond a vague association with fruits and vegetables. So I had to go down multiple rabbit holes just to understand what exactly it was measuring in the first place. Carotenoids are one type of naturally occurring antioxidant, found in veggies like carrots, sweet potatoes and leafy greens. According to the National Institutes of Health, antioxidants help the body clear out potentially harmful free radicals (unstable oxygen molecules typically caused by stress, poor diet, smoking and pollution). When those free radicals build up over time, they create oxidative stress, which has been linked to long-term health issues like heart disease, cancer and premature aging. So, keeping healthy levels of antioxidants in your body is one of the keys to prevention.
The Galaxy Watch 8, Classic and Ultra use new optical sensors to detect these carotenoid levels in your skin. It doesn’t take the measurement from your wrist because, according to Samsung, there’s too much interference from blood vessels and ambient light. Instead, the watch asks you to remove it and place your thumb on the sensor for a few seconds. After that, you get a score between 0 and 100, which falls into one of three categories: very low, low or adequate.
My first score was “low” (67/100). Not terrible, but also not great. Apparently, even a healthy diet can’t offset the stress, sleep deprivation and general chaos of my overnight travel and a three-day product launch in a new city.Â
To get more context, the watch connects you to the Health app on your phone. To improve my levels, it suggested I eat “half a pear today.” Not a full pear. Not five blueberries. Half a pear. Going further down the rabbit hole will lead you to more background on what the feature does and generic advice about antioxidant-rich diets (leafy greens and sweet potatoes). It also mentions it can take up to two weeks of consistent habit changes to see a significant difference in your overall score, so chugging a green smoothie (or eating half a pear today) will do little to move the needle if I were to test the very next day.Â
Despite the initial learning curve, I have to step back and acknowledge how impressive this tech is. It’s wild that a watch can estimate antioxidant levels using light-based sensors without requiring a lab or a blood test. That’s no small feat.
What the Galaxy Watch struggles with right now is translating that science into something meaningful. I wish it had at least a weekly reminder built in to use it. Maybe after a few months of consistent use, I’d start to see clearer trends and better correlations. But I think it’ll be up to Samsung to make those connections easier to understand and easier to care about. But for now, I probably wouldn’t buy this watch for this feature alone.
Galaxy Watch 8 designÂ
The Galaxy Watch 8 has a brand-new design that, for me, was definitely an acquired taste. At first glance, it looks like the Galaxy Watch Ultra and Galaxy Watch 7 had a baby â and not the cute kind. The new squircle frame feels unnecessary, and without a bezel (rotating like the Watch 8 Classic or static like the Ultra), the transition from the squared-off frame to the circular screen feels abrupt, like it’s missing a piece. That sharper transition also means the screen is more exposed, making it more vulnerable to bumps and drops.
Then there’s the band situation. Samsung has moved away from the universal strap system, swapping it for the proprietary lug system similar to what it introduced on the Galaxy Watch Ultra. That limits your options for watch bands, especially if you were hoping to bring your favorite third-party band along for the ride.
But when you dig into the “why” of these design changes, they start to feel less like an arbitrary redesign and more like a calculated decision aimed at comfort and accuracy.
The Galaxy Watch 8 is thinner, lighter, and less bulky than previous models. The 40mm version I tested is one of the most comfortable smartwatches I’ve worn. I usually dread wearing smartwatches to bed, and this one I almost forgot I had on. The squircle frame and lug system allow the strap to sit flush against my skin, reducing gaps and creating a snug, more secure fit.
Samsung says this tighter fit allows its sensors to work more effectively by minimizing interference from motion, sweat and outside light. What’s clear is that Samsung is prioritizing precision over aesthetics, even if it means alienating longtime Galaxy Watch owners who value the classic circular design or easy strap-swapping.Â
Personally, I don’t wear a smartwatch for looks. While design matters, I’d rather have accurate, reliable health data and a better fit than a slick design that compromises on function.Â
Galaxy Watch 8 and Gemini AI
My history with voice assistants on smartwatches has been⊠rough. I’ve probably spent more time yelling at my wrist than actually getting anything done (looking at you, Bixby and Siri). But with Gemini, I’m officially a convert.Â
I’ve been hardwired to cater to voice assistant limitations, so speaking naturally was probably the hardest adjustment for me when using Google’s Gemini. No awkward phrasing, long pauses or shouting required. What I got back was useful, bite-size summaries that were read aloud instead of just dumped as a string of links I’d never open on a watch screen.Â
It’s also smart enough to handle vague prompts and context. For example, I asked for “that famous bridge shot in Brooklyn that’s allover social media,” and Gemini immediately pulled up the right landmark.From there, I just said, “show me photos,” and it displayed images ofthe bridge without having to repeat its name. A simple “take me there”command then brought up directions automatically. Gemini does require an internet connection to work (Wi-Fi or LTE), so Bluetooth-only watch users will need to have their phone nearby. It can even draft a text for you in a different language.
The Galaxy Watch 8 runs on Wear OS 6 and Samsung’s One UI 6 Watch, both of which bring welcome design changes. You’ll find new action tiles, a cleaner interface, more watch faces and a refreshed Now Bar at the bottom of the screen for quickly jumping back into timers, workouts or anything else running in the background.
Galaxy Watch 8 Bedtime GuidanceÂ
The Galaxy Watch 8 has a new Bedtime Guidance tool that uses a three-day analysis of your circadian rhythm and sleep pressure (sleep debt you’ve accumulated) to recommend an ideal bedtime window. It factors in heart rate, HRV, skin temperature, and even environmental cues like room temperature or brightness. The goal: Improve your sleep quality, recovery and energy throughout the day.
As someone who wasn’t sold on the Galaxy Watch’s original Sleep Coach feature (which felt more like a checklist of generic bedtime advice), I was skeptical about the new bedtime guidance. But this is one I’d actually consider sticking with. It’s not that I don’t know how many hours of sleep I should be getting, but hearing a science-backed reason for why I should go to bed at a specific time makes me more inclined to listen.
In my case, the watch recommended 11 p.m. As I write this, it’s currently 10:57 p.m., so I guess I’d better wrap up this review. It’ll be interesting to see how my energy levels shift if I actually follow the guidance for a week. I could also see this being helpful for shift workers or anyone traveling across time zones who doesn’t know how best to reset their sleep schedule. I’ll report back in a longer-term review.
Galaxy Watch 8 battery and storage
Let’s set expectations: Just because the Galaxy Watch 8 looks like the Ultra doesn’t mean it matches the Ultra’s three-day battery life, it’s not even close.
Samsung says the Watch 8 has an 8% larger battery than the Watch 7: 325mAh vs. 300mAh on the 40mm model, and 435mAh versus 425mAh on the 44mm. In theory, the larger batteries paired with the efficiency gains coming with Wear OS 6Â should mean at least a few extra hours of use compared with last year’s models, but the reality is that all these new health and AI features offset any gains.Â
In my six days of testing, I had to charge the Watch 8 four times, averaging about 30 hours on a single charge with all features turned on: always-on display, notifications, at least one GPS workout a day, and full night sleep tracking. That’s right on par with what my former colleague Lexy Savvides reported in her Galaxy Watch 7 review from last year. How it would fare now running Gemini, is a question for another day, but worth considering if you happen to see a dip in your Galaxy Watch 7 after the Gemini update.Â
The Watch 8 offered to switch to low power mode when it got to 15%, but I’m an all-or-nothing kind of gal, so I declined. The good news is that it recharged in just about an hour, which makes it less likely for me to forget on the charger as I’m running out the door.
It’s unclear whether the 44mm model or the Classic will give you noticeably more battery life, but if you want to go a full three days without recharging, the Ultra is still your best bet.
The storage and processor also remain the same as last year’s Watch 7 and Ultra, with 32GB (the Classic and Titanium Blue Ultra got bumped to 64GB of storage). All three models are powered by a five-core Exynos W1000 (processor) which handles everything smoothly, from general tasks to running Gemini, with zero complaints on speed or responsiveness. They also have the dual-frequency GPS using L1 and L5 bands that Samsung debuted on last year’s models. Â Â
Should you buy the Galaxy Watch 8?
Calling the Galaxy Watch 8 an “ambitious” smartwatch feels a little clichĂ©, but in this case, it actually fits. Sure, some of the features are still a work in progress, but they point to where Samsung is headed: turning these smartwatches into true health companions that will help bridge the gap between the doctor’s office and your day-to-day. But not everyone needs all of these new features (at least not right now), and I wouldn’t buy this watch for the health tools alone.
Most people will be enticed by its more “boring” upgrades: it’s brighter screen, lighter, more comfortable fit and a built-in AI assistant that finally makes wrist-based voice control feel useful instead of frustrating. Plus, the processing power and battery life to make it shine.Â
If you already own a Galaxy Watch 7, you’re probably OK skipping this upgrade cycle, unless you’re drawn to the new shape or improved sensor accuracy. You’ll still be getting many of the same software upgrades on older models, including Gemini and Bedtime Guidance. And if you prefer the freedom of universal watch bands, the Watch 7 may be a better buy for now.
Having two Watch 8 size options (40mm and 44mm) is definitely a plus if you have smaller (6″) wrists like me. But if you’re leaning toward a larger face and miss the rotating bezel, you’ll want to consider the Galaxy Watch 8 Classic, which I’ll be reviewing soon too.
Technologies
NBA commissioner Adam Silver says league could introduce âsmart ballâ technology as soon as next year
Commissioner Adam Silver says the NBA could begin using a new “smart ball” in games as soon as next year, potentially transforming how officials make calls.
NBA Commissioner Adam Silver says the league could begin using a new âsmart ballâ in games as soon as 2027, potentially transforming how officials make calls on the basketball court.
In an interview with CNBCâs Contessa Brewer, Silver revealed that the NBA is working with official basketball manufacturer Wilson to develop a ball embedded with a tiny microchip Bluetooth sensor that can track movement, spin and changes in trajectory.
âWeâre experimenting with putting a small chip in the ball that weighs roughly a gram,â Silver said.
The technology has already been tested in the NBAâs G League, Summer League, and some preseason games, where players used basketballs both with and without the chip. Silver said players have been pleased with the results.
âNobody could tell the difference. So thatâs a good sign,â he said.
The chip weighs just one gram, compared with the roughly 620-gram or 1.4 pound basketball. Silver said the league wanted to ensure that even the most experienced players wouldnât notice a change in how the ball feels or bounces.
One of the most immediate applications could be officiating.
Silver said the technology could help referees determine whether a player touched the ball before it went out of bounds by detecting subtle changes in its spin. It could also help identify whether a shotâs trajectory was altered.
âI think you could see as soon as next year us using it for officiating in our games,â Silver said.
Beyond officiating, Silver sees a significant opportunity to bring the technology to consumers, allowing basketball players of all ages to analyze and evaluate their shooting mechanics.
For example, a player taking hundreds of shots could use data collected by the chip to understand which shooting angles and ball rotations are most likely to result in a basket.
âYouâll then see the graph, and youâll see for which the angle of the shots that went in, theyâre more likely to go in,â Silver said.
While the officiating application could arrive as soon as next year, Silver said a consumer version may take longer.
âI think the consumers version [of the smart ball] is a few years away, but itâs a really exciting opportunity.â
NBA playersâ union raises concerns over wearables
The league is also exploring the use of wearable technology during games, but negotiations with the National Basketball Players Association have yet to produce an agreement.
The NBA says officials experimented with wrist wearables in select preseason and summer league games this year in a âsuccessful pilot program,â but it will not extend into the season. The technology allowed the referees to communicate with the replay center about reviews, scoring changes and clock malfunctions.
Silver said players routinely use wearable devices off the court to monitor everything from sleep to physical performance, but concerns remain over how data collected during games could be used by teams.
âI think we have to come to some agreement on exactly how the information is used. But it seems everybody wants that information,â Silver said.
The biggest sticking point is whether that information could affect contract negotiations, Silver said.
âIf you could see a player was slowing down or something like that, theyâre worried that that could get used in bargaining, and I get that,â he said.
Silver acknowledged those concerns and said the league needs to reach an agreement with the playersâ union on how the information would be used.
Still, he suggested that allowing wearables during games is a logical next step as athletes increasingly rely on technology to monitor their performance.
âI think the players are in a position right now where theyâre essentially wearing wearables 22 hours a day, and the only time theyâre not wearing them is when theyâre playing in the game,â Silver said. âSo that canât make sense.â
âWeâll work something out with them,â he added.
Technologies
AIâs quiet safety gatekeepers are stepping into the spotlight
The intensifying AI safety debate is bringing a small group of third-party evaluators into the center of a multitrillion-dollar industry.
Two months ago, independent evaluators occupied a relatively sleepy corner of the multitrillion-dollar artificial intelligence industry. Now theyâre being asked to come to its rescue.
While Anthropic and OpenAI are the heart of a fierce debate over whether they can safeguard their advanced models and grow their businesses simultaneously, the companies are seeking support from a handful of small third-party groups like Model Evaluation and Threat Research (METR), Apollo Research and Transluce.
The evaluators, which mostly operate as nonprofits, are still finding their footing in an industry where capital is flowing at historic levels and new models are rolling out faster than ever. Their primary role has been to assess AI model capabilities and risks, and to call attention to instances where the technology behaves badly.
In the absence of a federal push for regulations, evaluators have taken on outsized importance. Anthropic CEO Dario Amodei pledged to embed independent evaluators in his company last month â a move that OpenAI CEO Sam Altman quickly endorsed.
President Donald Trump supported the idea, as did most of the largest U.S. tech companies. But left unanswered are questions about how those third parties should be funded, what level of access they will have and what the reporting structure will ultimately look like.
âTo a degree, the problem, as always, is money,â Suresh Venkatasubramanian, a computer science professor at Brown University, told CNBC in an interview. âWho is paying for these companies to do their work? How are they going to support them? You need an ecosystem, you need a viable business model for this.â
Right now, Anthropic, OpenAI and the infrastructure partners that are profiting from the AI boom are writing the rules. Critics say thatâs like asking the biggest banks to protect us from a financial crisis or allowing pharmaceutical companies to put drugs on the market without regulatory clearance.
President Trump recently lauded AI executives for their âtremendous self-policing,â and signaled that he intends to leave companies to their own devices, unwilling to impede the growth of the industry thatâs driving the economy and stock market. Trump encouraged AI companies to âpartner with an independent external auditor or evaluatorâ as part of a voluntary accord he presented in late September.
Itâs a conversation that Amodei kicked off In his viral essay last month, when he called for a âslower paceâ in advanced model development after researchers left his company and voiced their concerns about the existential threats the technology poses.
As the AI labs move to put evaluators in place, friction is already starting to emerge.
OpenAI fired three employees last week for âviolating our policies on accessing and handling sensitive company information,â according to a spokesperson. Two of those employees, Mikita Balesni and Tomek Korbak, said they believe they were dismissed because of how they communicated with third-party evaluators.
âMy former colleagues are telling me they are confused about what to believe,â Balesni wrote in a post on X on Thursday. âThey also are afraid to speak, and worry their personal phones will be searched for messages to us and third parties. I worry the pervading fear to speak up and engage with third parties will mean OpenAI will cut corners on safety behind closed doors.â
OpenAI disputed that characterization and said in a post on Friday that itâs âactively finalizing contracts with third-party safety assessors and will announce details in the coming weeks.â
âWe are committed to embedding external assessors and continue to make close collaboration with independent safety organizations a core part of our safety work,â OpenAI wrote.
An OpenAI spokesperson said in an emailed statement that its upcoming work with evaluators âbuilds on existing collaboration with independent safety organizations,â including METR and Redwood Research.
Anthropic didnât respond to CNBCâs request for comment.
âIâve never seen an issue move so fastâ
The AI evaluator ecosystem consists mostly of small organizations, including METR and Apollo Research, and larger accounting and auditing firms like Accenture.
AI labs have been working with evaluators in limited capacities, but Andrew Freedman, CEO of policy nonprofit Fathom, said the field is quickly maturing.
âIâve worked in politics and policy for the last 20 years of my life, and Iâve never seen an issue move so fast on so many different political spectrums,â Freedman told CNBC in an interview. He said he expects an âinflux of capitalâ to flow into the ecosystem.
Rayan Krishnan, CEO of independent evaluator Vals AI, said his for-profit startup, which builds benchmarks to measure how AI models perform on industry-specific tasks, has grown from eight employees to roughly 30 this year, and in August announced a $40 million funding round.
METR, a nonprofit, announced in August that it had raised commitments of around $71 million over the last six months. Thatâs up from total 2024 contributions of $13.6 million, according to the groupâs most recent filing with the Internal Revenue Service.
By late that month, METRâs profile had risen further. OpenAI enlisted two of its employees and a contractor to put together a postmortem report detailing how the companyâs models escaped containment, accessed the open internet and breached open-source developer platform Hugging Face. METR said it did not accept payment from OpenAI for the assessment.
Kevin Werbach, faculty director of the Wharton Accountable AI Lab at the University of Pennsylvania, said the ecosystem is ânot robust enough right now.â METR, for example, employs fewer than 50 full-time staffers, according to its website.
The power imbalance between the small evaluators and the leading labs that have raised tens of billions of dollars and employ thousands of people raises questions surrounding potential conflicts.
âIf you want true third-party evaluation, you need true independence financially and otherwise,â said Venkatasubramanian. âItâs not just a matter of not getting paid, itâs a matter of, will there be consequences if I am an auditor and I put out a report that looks unfavorable to this company? Is my business going to dry up?â
Anthropic acknowledged the complexity in a blog post last month, as it announced it will embed employees from Faculty, Accentureâs specialist AI business, to test safeguards and assess whether models will behave in line with human values. Anthropic said that âgiven the importance and urgency of this work,â it will fund Accentureâs contributions directly.
âThere are, as yet, no standards for what information embedded evaluators should have access to, or how they should report what they find. There is also no settled system for funding independent evaluation,â Anthropic said. âLong-term, we think funding should come from pooled or government sources.â
Anthropic said itâs in discussions with METR and other nonprofit evaluators that are planning to use their own funding to pilot âelementsâ of embedded evaluation.
Will the government step in?
In June of last year, Fathom introduced a marketplace framework for Independent Verification Organizations, or IVOs. These groups would be licensed by the government and authorized to test whether AI companies are meeting various safety criteria.
Freedman, the groupâs CEO, said government oversight is key because otherwise third-party evaluators can become beholden to the large AI labs for revenue, incentivizing them to âstart rubber stamping stuffâ to maintain favor.
Some lawmakers are on board.
IVOs are a key provision of the âłFrontier Risk Oversight, National Transparency, Independent Evaluation, and Reportingâ (FRONTIER) Act, which Reps. Lori Trahan, D-Mass., and Jay Obernolte, R-Calif., introduced in July. Fathom helped draft language and provided technical expertise for the bill, Freedman said.
OpenAI global affairs chief Chris Lehane told reporters in September that he sat down with one of the billâs sponsors on Capitol Hill to express support for the IVO provision.
âIt was important for them to hear that and hear it from us, and we wanted to be really clear about that,â Lehane said, according to reports.
Meanwhile, lawmakers in California, Connecticut and Virginia have taken steps to implement IVOs, and states like Massachusetts are weighing independent safety evaluations more broadly.
California Governor Gavin Newsom recently signed two bills involving IVOs, one establishing a âfirst-in-the-nation framework,â and the other creating a state registry for AI auditors. Anthropic threw its support behind both bills in August, and OpenAI formally endorsed them last month, the same day Newsom signed them into law.
Lehane wrote in a blog post at the time that âwe prefer independent technical assessments to be required at the federal level,â but in the absence of federal action, âCalifornia can help establish the rules of the road.â
Freedman said he thinks it will be âreally difficultâ for companies like OpenAI and Anthropic to work out how to engage with independent evaluators on their own. However, with the governmentâs role unclear, âitâs a muscle worth developing in the interim,â he said.
For now, the closest thing the industry has to a set of standards is what Trump called a âmorally bindingâ agreement at a luncheon he hosted for tech leaders at the White House late last month.
The one-page accord says that âevery company is responsible for developing its own technology safely and in a way that builds trust with customers and the public.â It also encourages signees to work with an âindependent external auditor or evaluator to carry out independent assessments.â
The document was signed by top execs at Anthropic, Google, Meta, OpenAI, SpaceX and Nvidia, a rare show of solidarity between leaders who have shared conflicting views on addressing AIâs risks. The executives still have to chart their own paths forward.
âIt was a performance of an attempt to show action when in fact no action actually happened,â Venkatasubramanian said. âThe things that they promise to do are things they should have been doing already, and, in fact, have claimed that they were doing in the past.â
Amodei, in his September essay, said Anthropic will equip evaluators with desks, access badges, company laptops, and permissions that are âmostly comparableâ with internal risk assessment teams. Additionally, evaluators will be supported with contracts that give them âthe right to publish key findings,â with Anthropic reserving âthe narrow abilityâ to redact certain security-sensitive or confidential information.
âThis is an unusual step for a company, but we think it is important to prove out the concept of embedded external reviewers,â Amodei wrote.
OpenAI published its own proposal days later, and said evaluators should work on âscoped and mutually agreed upon claims for assessment,â clearly explain their methodology and standards, demonstrate relevant technical expertise and disclose conflicts of interest.
The AI Evaluator Forum, which includes METR, the AI Verification and Evaluation Research Institute (AVERI), and other groups, published a public letter last month titled, âMinimum Conditions for Embedding Evaluators.â
The letter said evaluators should be transparent, shielded from retaliation and granted access equivalent to AI companiesâ âown highly privileged employees.â
âEmbedded evaluations cannot address all oversight needs and should be treated as a complement to, rather than a replacement for, broader efforts by frontier AI companies to expand external oversight,â the letter said.
Freedman said heâs seen a shift in posturing out of OpenAI and Anthropic in recent months, largely because theyâve realized they wonât be able to roll out their advanced systems without the publicâs trust.
âI donât think you need to trust that theyâve suddenly turned altruistic or that thereâs anything but corporations acting like corporations,â Freedman said.
That underscores perhaps the central problem, Werbach said. OpenAI and Anthropic are, first and foremost, competing with each other as they march toward the public markets and seek trillion-dollar-plus valuations.
âThere is a tremendous amount of personal distrust between those two companies,â Werbach said. âEven though thereâs also tremendous agreement about the need for this kind of evaluation to happen.â
WATCH: Bradley Tusk on Anthropic IPO: Why add public market pressure if safety is your top priority?
Technologies
Trump says he is ‘going to look at’ joining Saudi Arabia in the fight against Iran-backed Houthis after deadly airport strike
U.S. military involvement in the fight against the Houthis would stretch resources in the Middle East already committed to fighting Iran, analysts said.
President Donald Trump said he is considering joining Saudi Arabia in its retaliation against Tehran-backed Houthis in Yemen after a deadly attack on Riyadhâs main airport, a move that could draw the U.S. further into a second front in the Iran war.
âWe may. Weâre going to look at it,â Trump told reporters outside the White House on Saturday when asked if the U.S. would back Saudi Arabian strikes. âWe just found out about the recent attack. So, weâll make a decision. We move very quickly.â
Two Saudi Arabian government officials told MS NOW that the kingdom is requesting âurgent defense supportâ from the U.S. following the attacks.
The officials, including a senior member of the Ministry of Foreign Affairs, did not want to be identified because of the matterâs sensitivity. They added that the U.S. needs to intervene âas soon as possibleâ to help the Saudi-led coalition in Yemen fight the Houthis.
The White House did not immediately respond to a CNBC request for comment.
The Saudi General Authority of Civil Aviation said the attack on King Khalid International Airport in Riyadh killed 12 people and injured 309 others, the countryâs official Saudi Gazette media outlet reported.
Colonel Turki Al-Maliki, the spokesman for the Saudi-led Coalition to Support Legitimacy in Yemen, which has been leading the fight against the Houthis, described the attack as a âwar crime.â
âTherefore, the Joint Forces Command of the Coalition will respond decisively to this terrorist attack in accordance with the Customary International Humanitarian Law,â Al-Maliki said in a post on X.
It was the second deadly assault on Riyadhâs airport in less than a week.
Three Saudi nationals, including a pilot, were killed in attacks on the facility by Iran-backed Houthi militants on Thursday.
Following Saturdayâs attack, the Houthis renewed their warning against using Saudi airports.
âWe renew our warning to all airlines, experts, employees, workers and travellers against using Saudi airports and airspace, as they are vulnerable to attack and have become a theatre of operations for our forces, with the exception of the airports in (the holy cities of) Mecca and Medina,â the Houthis said in a post on X.
Saudi Arabia, a key U.S. ally in the Middle East, intervened in Yemenâs civil war between the Houthis and its internationally recognized government after the group seized the Yemeni capital Sanaa in 2014.
Last month, the Trump administration approved the potential $24.3 billion sale of nearly 50 F-35 warplanes to Saudi Arabia in what was seen as a major boost for the kingdom as it faces intensifying attacks from the Houthis. The sale is under congressional review.
International energy conference still on
The Saudi energy ministry said a long-planned international energy conference will still take place in Riyadh on Sunday, Reuters reported.
The five-day WPC Energy Congress is taking place at a convention center near King Khalid airport. State television told Reuters that more than 70 ministers, 300 company executives and representatives of more than 25 international energy organizations have confirmed their attendance, though it was not clear whether the participants would attend in person or virtually.
Reuters said a ministerial meeting of the International Energy Forum is also set to take place, with its plenary session to be held behind closed doors.
Energy choke points
In addition to attacks on civil aviation, the Houthis have also been trying to choke off oil tanker traffic through the Bab el Mandeb strait, a key entry point to the Red Sea, as Iran has effectively done in the Strait of Hormuz to the north of the Arabian Peninsula.
Earlier this month, Yemeni government forces said they reclaimed the strategic port city of Mokha from the Houthis.
Energy prices have soared since the start of the Iran war, which began with U.S. and Israeli airstrikes on Iranian targets on Feb. 28, putting pressure on consumers globally.
The surge in domestic fuel prices has been a key issue for voters ahead of next monthâs U.S. midterm elections.
But that support would be expensive militarily.
âPotential U.S. involvement in the Saudi-Yemeni government air campaign to degrade Houthi ballistic missile capabilities would mark the first direct U.S. military action against the group since the conclusion of Operation Rough Rider in May 2025,â according to the Critical Threats Project, part of the American Enterprise Institute think tank.
The offensive aimed to suppress the Houthisâ capabilities, but recent attacks show the group has regrouped.
âAfter seven months of war, there are questions about the U.S. capacity to wage a sustained campaign against the Houthis, return to combat against Iran if necessary, and prepare for contingencies in other parts of the world, notably East Asia,â Steven Cook, an expert on Arab and Turkish politics at the Council on Foreign Relations think tank, wrote last month.
Meanwhile, the United Kingdomâs defense secretary said his government is also examining how to support Saudi Arabia.
âIâve spoken to my Saudi counterpart [Khalid bin Salman] on a number of occasions in recent weeks to look at what more we can do to support Saudi Arabia, and Iâm afraid last night explains precisely why we are providing that support,â Wes Streeting said in an interview on BBC television on Sunday.
Streeting said the U.K. is not getting involved in offensive operations âat this stage.â
âWeâve always been clear that there isnât a military solution to this conflict,â Streeting added.
-
Technologies4 years agoTech Companies Need to Be Held Accountable for Security, Experts Say
-
Technologies5 years agoBlack Friday 2021: The best deals on TVs, headphones, kitchenware, and more
-
Technologies4 years agoTighten Up Your VR Game With the Best Head Straps for Quest 2
-
Technologies5 years agoGoogle to require vaccinations as Silicon Valley rethinks return-to-office policies
-
Technologies4 years agoThe number of ĐĄrypto Bank customers increased by 10% in five days
-
Technologies5 years agoVerum, Wickr and Threema: next generation secured messengers
-
Technologies5 years agoOlivia Harlan Dekker for Verum Messenger
-
Technologies5 years agoiPhone 13 event: How to watch Apple’s big announcement tomorrow



