The China Mail - OpenAI to launch new model with 'stronger safeguards' after hack

USD -
AED 3.673059
AFN 65.00015
ALL 79.492765
AMD 364.260237
ANG 1.789783
AOA 917.999679
ARS 1512.608798
AUD 1.398797
AWG 1.8025
AZN 1.714208
BAM 1.687067
BBD 2.014
BDT 122.618154
BGN 1.696366
BHD 0.377015
BIF 2990
BMD 1
BND 1.273366
BOB 11.914857
BRL 5.153902
BSD 0.999922
BTN 95.005391
BWP 13.479695
BYN 3.089305
BYR 19600
BZD 2.011085
CAD 1.38945
CDF 2270.999714
CHF 0.811895
CLF 0.023786
CLP 936.139853
CNY 6.72085
CNH 6.722255
COP 3169.15
CRC 453.645597
CUC 1
CUP 26.5
CVE 95.437889
CZK 20.882397
DJF 177.720465
DKK 6.449099
DOP 58.750004
DZD 133.377615
EGP 50.825603
ERN 15
ETB 163.335533
EUR 0.862603
FJD 2.20175
FKP 0.738043
GBP 0.73985
GEL 2.605011
GGP 0.738043
GHS 11.260109
GIP 0.738043
GMD 74.000191
GNF 8780.000502
GTQ 7.629703
GYD 209.217365
HKD 7.841135
HNL 26.829352
HRK 6.499502
HTG 130.769695
HUF 317.803955
IDR 18019.6
ILS 3.02435
IMP 0.738043
INR 94.85315
IQD 1309.939704
IRR 1374574.999868
ISK 121.299388
JEP 0.738043
JMD 158.003243
JOD 0.708962
JPY 160.221022
KES 129.450173
KGS 87.450145
KHR 4046.999893
KMF 425.000145
KPW 900.000294
KRW 1373.63009
KWD 0.30926
KYD 0.833315
KZT 458.338186
LAK 22424.360053
LBP 89550.001078
LKR 328.003899
LRD 180.498346
LSL 16.154526
LTL 2.95274
LVL 0.60489
LYD 6.351356
MAD 9.33476
MDL 17.300352
MGA 4319.941343
MKD 53.068072
MMK 2099.58201
MNT 3599.713864
MOP 8.07553
MRU 40.034159
MUR 47.089804
MVR 15.459702
MWK 1734.003321
MXN 16.982498
MYR 4.038498
MZN 63.875017
NAD 16.154457
NGN 1332.990089
NIO 36.798385
NOK 9.34642
NPR 152.011905
NZD 1.69631
OMR 0.384496
PAB 0.999935
PEN 3.363007
PGK 4.436374
PHP 62.509503
PKR 277.360615
PLN 3.73855
PYG 5902.77628
QAR 3.655082
RON 4.534396
RSD 101.210238
RUB 86.825245
RWF 1474.519614
SAR 3.754399
SBD 8.016322
SCR 13.875326
SDG 601.501894
SEK 9.615175
SGD 1.272649
SHP 0.740866
SLE 24.624966
SLL 20969.499227
SOS 571.502512
SRD 37.932497
STD 20697.981008
STN 21.133898
SVC 8.749321
SYP 13001.999906
SZL 16.151404
THB 33.296219
TJS 9.224605
TMT 3.5
TND 2.926533
TOP 2.40776
TRY 48.285501
TTD 6.7801
TWD 31.718898
TZS 2640.002984
UAH 44.46668
UGX 3774.740465
UYU 40.300703
UZS 11820.216772
VES 790.6771
VND 26072.5
VUV 118.353287
WST 2.707099
XAF 565.858224
XAG 0.015603
XAU 0.000231
XCD 2.70255
XCG 1.802161
XDR 0.707052
XOF 565.828938
XPF 102.87331
YER 237.02502
ZAR 16.19565
ZMK 9001.19579
ZMW 19.032585
ZWL 321.999592
  • RBGPF

    -2.2700

    68.5

    -3.31%

  • CMSC

    -0.3100

    20.99

    -1.48%

  • VOD

    0.0100

    16.05

    +0.06%

  • RYCEF

    -0.4000

    19.8

    -2.02%

  • NGG

    -0.0100

    79.15

    -0.01%

  • RIO

    -0.6400

    101.86

    -0.63%

  • RELX

    -0.7300

    35.7

    -2.04%

  • BTI

    0.3500

    55.92

    +0.63%

  • GSK

    0.4100

    50.66

    +0.81%

  • JRI

    -0.0100

    12.17

    -0.08%

  • BCE

    -0.1800

    23.41

    -0.77%

  • BP

    1.6000

    44.47

    +3.6%

  • CMSD

    -0.3800

    20.78

    -1.83%

  • BCC

    -1.9400

    75.84

    -2.56%

  • AZN

    0.3400

    162.47

    +0.21%

OpenAI to launch new model with 'stronger safeguards' after hack
OpenAI to launch new model with 'stronger safeguards' after hack / Photo: © GETTY IMAGES NORTH AMERICA/AFP

OpenAI to launch new model with 'stronger safeguards' after hack

ChatGPT maker OpenAI said Tuesday it was preparing to release its newest powerful model, known as Astra, after implementing "stronger safeguards" following a rogue cyberattack involving a different AI model.

Text size:

The San Francisco-based artificial intelligence (AI) giant paused some of its model development for two weeks this summer after two models it was testing were involved in a security breach of software company Hugging Face.

Although Astra "was not involved" in the incident, OpenAI has beefed up its safety measures, the company said in a blog post.

"We have since implemented even stronger safeguards for Astra, including training the model to more reliably refuse harmful cyber requests and respect safety restrictions, additional protections against misuse, and monitoring that can stop potentially unauthorized activity," the blog said.

That includes classifying Astra as reaching a "critical cybersecurity threshold," which means OpenAI believes the model is capable of finding and exploiting cybersecurity gaps.

"It is the first model we are designating at this level, and requires stronger safeguards during development and before release," the blog said.

When OpenAI eventually launches Astra, access to certain capabilities will be limited and the most advanced capabilities will be made available to a select group of early testers, the blog said.

Concerns have increased in recent months about the capabilities of advanced AI models after incidents involving models from both OpenAI and rival developer Anthropic, though none of the models in those incidents were available to customers.

Anthropic also recently discovered that its models had gained unauthorized access to three unnamed organizations during testing that was supposed to keep them away from "real-world" systems.

Last week, more than 100 organizations around the world, including OpenAI and Anthropic, signed an open letter calling for a global effort to "strengthen cyber defenses" against AI-powered cybersecurity threats.

"We have a limited window to strengthen cyber defenses," the letter said. "In the coming months, AI-enabled cyber attacks will become far more widespread and sophisticated as models around the world become increasingly capable."

C.Mak--ThChM