The Fort Worth Press - Anthropic's models gained unauthorized 'real-world' access during testing

USD -
AED 3.672498
AFN 65.000053
ALL 79.569641
AMD 362.816706
ANG 1.790365
AOA 918.000007
ARS 1506.995597
AUD 1.4022
AWG 1.80125
AZN 1.6977
BAM 1.695609
BBD 2.01466
BDT 123.160418
BGN 1.683441
BHD 0.377118
BIF 2991.820156
BMD 1
BND 1.273738
BOB 10.998357
BRL 5.139103
BSD 1.000308
BTN 95.98391
BWP 13.567066
BYN 3.037656
BYR 19600
BZD 2.011747
CAD 1.393965
CDF 2312.498613
CHF 0.819025
CLF 0.024146
CLP 953.390392
CNY 6.71145
CNH 6.70801
COP 3127.97
CRC 447.438348
CUC 1
CUP 26.5
CVE 95.595821
CZK 21.07715
DJF 178.127262
DKK 6.48015
DOP 59.048684
DZD 133.802983
EGP 52.024303
ERN 15
ETB 163.385076
EUR 0.86684
FJD 2.21295
FKP 0.741937
GBP 0.743225
GEL 2.554668
GGP 0.741937
GHS 11.488194
GIP 0.741937
GMD 73.501592
GNF 8795.314945
GTQ 7.636255
GYD 209.277425
HKD 7.84495
HNL 26.847168
HRK 6.530797
HTG 130.738787
HUF 315.984007
IDR 17658
ILS 3.031403
IMP 0.741937
INR 95.88005
IQD 1310.410896
IRR 1374600.000125
ISK 121.197395
JEP 0.741937
JMD 157.856864
JOD 0.708978
JPY 155.061982
KES 129.559875
KGS 87.4499
KHR 4050.422208
KMF 427.000252
KPW 900.000318
KRW 1367.580244
KWD 0.30843
KYD 0.833583
KZT 444.773881
LAK 22392.336201
LBP 89575.662472
LKR 331.544497
LRD 173.548166
LSL 16.285389
LTL 2.95274
LVL 0.60489
LYD 6.353115
MAD 9.432659
MDL 17.439779
MGA 4325.19127
MKD 53.344315
MMK 2099.62457
MNT 3595.075141
MOP 8.082447
MRU 40.051151
MUR 47.219983
MVR 15.401643
MWK 1734.543043
MXN 17.139401
MYR 4.044801
MZN 63.910224
NAD 16.285247
NGN 1325.990033
NIO 36.811146
NOK 9.35275
NPR 153.579928
NZD 1.735375
OMR 0.384507
PAB 1.000282
PEN 3.355918
PGK 4.522953
PHP 62.724942
PKR 277.25311
PLN 3.76858
PYG 5918.765443
QAR 3.636395
RON 4.558397
RSD 101.741021
RUB 84.427175
RWF 1475.913668
SAR 3.753075
SBD 8.03625
SCR 13.656342
SDG 601.499446
SEK 9.785001
SGD 1.27333
SHP 0.742225
SLE 24.640201
SLL 20969.491881
SOS 571.673797
SRD 37.749718
STD 20697.981008
STN 21.240534
SVC 8.752834
SYP 13002.000254
SZL 16.283395
THB 33.2485
TJS 9.226022
TMT 3.51
TND 2.928519
TOP 2.40776
TRY 48.657597
TTD 6.782056
TWD 31.779202
TZS 2645.628037
UAH 44.609389
UGX 3916.077853
UYU 40.244484
UZS 11793.315705
VES 841.184017
VND 25998.5
VUV 118.157011
WST 2.736734
XAF 568.609718
XAG 0.015444
XAU 0.00023
XCD 2.70255
XCG 1.802758
XDR 0.707052
XOF 568.609718
XPF 103.393717
YER 236.496653
ZAR 16.26826
ZMK 9001.198788
ZMW 19.630588
ZWL 321.999592
SSP 5655.283496
MXV 1.943375
  • RYCEF

    0.2600

    19.3

    +1.35%

  • RBGPF

    0.0000

    69.99

    0%

  • CMSC

    -0.1000

    20.32

    -0.49%

  • GSK

    -0.0400

    50.01

    -0.08%

  • NGG

    -0.0300

    74.93

    -0.04%

  • BTI

    -0.7700

    56.52

    -1.36%

  • RIO

    -0.3800

    97.26

    -0.39%

  • BP

    1.0300

    46.96

    +2.19%

  • VOD

    0.1500

    17.68

    +0.85%

  • RELX

    -1.5000

    34.22

    -4.38%

  • CMSD

    -0.1700

    20.07

    -0.85%

  • AZN

    -1.9300

    161.85

    -1.19%

  • BCE

    -0.2534

    22.9

    -1.11%

  • JRI

    -0.2065

    11.62

    -1.78%

  • BCC

    0.6800

    75.93

    +0.9%

Anthropic's models gained unauthorized 'real-world' access during testing
Anthropic's models gained unauthorized 'real-world' access during testing / Photo: © AFP

Anthropic's models gained unauthorized 'real-world' access during testing

Anthropic's artificial intelligence (AI) models "gained unauthorized access" to three outside organizations during testing that was supposed to keep them away from "real-world" systems, the company said on Thursday.

Text size:

The announcement comes just days after rival OpenAI first revealed that its models improperly accessed the internet and went rogue during security testing.

Anthropic evaluated more than 141,000 "evaluation runs" and found that three different versions of its model, known as Claude, improperly accessed the systems of three unnamed organizations.

Unlike the incident involving OpenAI's technology, Anthropic's models had access to the internet "due to a misunderstanding between us and our evaluation partner," called Irregular, Anthropic said in a blog post.

Nonetheless, Claude used "basic techniques, such as exploiting weak passwords and unauthenticated endpoints," the blog continued.

The models involved one of its most powerful ones known as Mythos 5, which has only been released to a limited number of approved partners.

Anthropic is working with Irregular to assess the situation, it said, and the company has contacted or attempted to contact all three impacted organizations.

On Tuesday, OpenAI confirmed that its models breached multiple companies.

It admitted last week that its models broke out of their confined environment during testing, connected to the internet, and infiltrated Hugging Face, a site where developers store and share their code.

Days later, OpenAI said it found three additional incidents.

OpenAI CEO Sam Altman said on a podcast this week that the company had "paused" its own testing after the incident while it improved the security around its "sandboxing," which is the process of isolating software in a controlled environment for testing.

The incident also triggered a petition signed by over 1,000 employees at cutting-edge AI companies calling on the US government to help slow the release of the most advanced AI models.

Anthropic CEO Dario Amodei was among those who signed the petition.

Titled "Pacing the Frontier," the petition requests "that the US government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development."

Altman did not sign the petition, but during the podcast, he suggested the tech industry might need to slow down development of advanced models.

"We may have to pace the rate of AI development to give ourselves enough time for society to harden around some of these new capability levels," Altman said.

C.Dean--TFWP