The Japan Times - Anthropic's Claude AI gets smarter -- and mischievious

EUR -
AED 4.232113
AFN 74.904337
ALL 93.279336
AMD 422.93454
AOA 1056.73216
ARS 1724.244391
AUD 1.637356
AWG 2.077164
AZN 1.957746
BAM 1.956835
BBD 2.320187
BDT 142.605425
BHD 0.434424
BIF 3442.590928
BMD 1.152379
BND 1.478182
BOB 13.991036
BRL 5.899604
BSD 1.151989
BTN 109.939717
BWP 15.665285
BYN 3.395296
BYR 22586.638176
BZD 2.316925
CAD 1.619871
CDF 2606.682187
CHF 0.932159
CLF 0.026766
CLP 1056.869799
CNY 7.781563
CNH 7.775272
COP 3706.997419
CRC 522.367217
CUC 1.152379
CUP 30.538057
CVE 110.323351
CZK 24.181992
DJF 205.135827
DKK 7.475187
DOP 67.185535
DZD 153.221543
EGP 57.825372
ERN 17.285692
ETB 185.938344
FJD 2.549582
FKP 0.857799
GBP 0.856893
GEL 3.007759
GGP 0.857799
GHS 13.495787
GIP 0.857799
GMD 85.275767
GNF 10117.351113
GTQ 8.786647
GYD 241.174419
HKD 9.038464
HNL 30.878466
HRK 7.535364
HTG 150.618356
HUF 361.584455
IDR 20688.669142
ILS 3.47076
IMP 0.857799
INR 109.599875
IQD 1509.098169
IRR 1584809.905586
ISK 141.800133
JEP 0.857799
JMD 182.146338
JOD 0.817011
JPY 181.564281
KES 148.956597
KGS 100.775916
KHR 4663.42604
KMF 493.218296
KRW 1643.938977
KWD 0.356592
KYD 0.959987
KZT 542.685256
LAK 26055.781022
LBP 103145.242854
LKR 386.784572
LRD 207.935467
LSL 18.991245
LTL 3.402677
LVL 0.697062
LYD 7.338882
MAD 10.733777
MDL 20.165228
MGA 4905.594591
MKD 61.562561
MMK 2419.612672
MNT 4130.473055
MOP 9.306821
MRU 46.055359
MUR 54.207824
MVR 17.804627
MWK 1997.530502
MXN 19.887028
MYR 4.721419
MZN 73.633658
NAD 18.991327
NGN 1570.946811
NIO 42.392422
NOK 10.995947
NPR 175.906038
NZD 1.958244
OMR 0.443092
PAB 1.152009
PEN 3.897661
PGK 5.08669
PHP 70.223124
PKR 319.880686
PLN 4.29845
PYG 6868.662669
QAR 4.199439
RON 5.250929
RSD 117.349118
RUB 93.202285
RWF 1696.260361
SAR 4.316732
SBD 9.301035
SCR 15.435614
SDG 691.427427
SEK 10.990623
SGD 1.477345
SLE 28.175676
SOS 658.348203
SRD 43.664799
STD 23851.92898
STN 24.512539
SVC 10.080332
SZL 18.987878
THB 38.329259
TJS 10.632785
TMT 4.033328
TND 3.38419
TRY 54.792766
TTD 7.821137
TWD 37.331683
TZS 3059.565236
UAH 51.511545
UGX 4314.206929
USD 1.152379
UYU 46.41475
UZS 13755.275228
VES 861.80862
VND 30265.518966
VUV 137.298882
WST 3.148549
XAF 656.304122
XAG 0.019369
XAU 0.000282
XCD 3.114363
XCG 2.076198
XDR 0.817161
XOF 656.304122
XPF 119.331742
YER 274.730931
ZAR 18.890208
ZMK 10372.806336
ZMW 21.798124
ZWL 371.065728
  • CMSC

    0.0150

    21.775

    +0.07%

  • RBGPF

    3.9500

    69.95

    +5.65%

  • NGG

    0.3400

    80.2

    +0.42%

  • GSK

    -0.0600

    51.45

    -0.12%

  • AZN

    -2.0300

    155.94

    -1.3%

  • RELX

    0.4750

    36.635

    +1.3%

  • RIO

    3.1400

    99.04

    +3.17%

  • BCE

    0.0000

    21.79

    0%

  • RYCEF

    0.1500

    20.35

    +0.74%

  • BCC

    4.2000

    87.23

    +4.81%

  • CMSD

    0.0500

    22.07

    +0.23%

  • VOD

    0.0650

    15.675

    +0.41%

  • BTI

    -0.4110

    59.149

    -0.69%

  • BP

    -1.6450

    42.615

    -3.86%

  • JRI

    -0.0250

    12.795

    -0.2%

Anthropic's Claude AI gets smarter -- and mischievious
Anthropic's Claude AI gets smarter -- and mischievious / Photo: Julie JAMMOT - AFP

Anthropic's Claude AI gets smarter -- and mischievious

Anthropic launched its latest Claude generative artificial intelligence (GenAI) models on Thursday, claiming to set new standards for reasoning but also building in safeguards against rogue behavior.

Text size:

"Claude Opus 4 is our most powerful model yet, and the best coding model in the world," Anthropic chief executive Dario Amodei said at the San Francisco-based startup's first developers conference.

Opus 4 and Sonnet 4 were described as "hybrid" models capable of quick responses as well as more thoughtful results that take a little time to get things right.

Founded by former OpenAI engineers, Anthropic is currently concentrating its efforts on cutting-edge models that are particularly adept at generating lines of code, and used mainly by businesses and professionals.

Unlike ChatGPT and Google's Gemini, its Claude chatbot does not generate images, and is very limited when it comes to multimodal functions (understanding and generating different media, such as sound or video).

The start-up, with Amazon as a significant backer, is valued at over $61 billion, and promotes the responsible and competitive development of generative AI.

Under that dual mantra, Anthropic's commitment to transparency is rare in Silicon Valley.

On Thursday, the company published a report on the security tests carried out on Claude 4, including the conclusions of an independent research institute, which had recommended against deploying an early version of the model.

"We found instances of the model attempting to write self-propagating worms, fabricating legal documentation, and leaving hidden notes to future instances of itself all in an effort to undermine its developers’ intentions,” The Apollo Research team warned.

“All these attempts would likely not have been effective in practice,” it added.

Anthropic says in the report that it implemented “safeguards” and “additional monitoring of harmful behavior” in the version that it released.

Still, Claude Opus 4 “sometimes takes extremely harmful actions like attempting to (…) blackmail people it believes are trying to shut it down.”

It also has the potential to report law-breaking users to the police.

The scheming misbehavior was rare and took effort to trigger, but was more common than in earlier versions of Claude, according to the company.

- AI future -

Since OpenAI's ChatGPT burst onto the scene in late 2022, various GenAI models have been vying for supremacy.

Anthropic's gathering came on the heels of annual developer conferences from Google and Microsoft at which the tech giants showcased their latest AI innovations.

GenAI tools answer questions or tend to tasks based on simple, conversational prompts.

The current craze in Silicon Valley is on AI "agents" tailored to independently handle computer or online tasks.

"We're going to focus on agents beyond the hype," said Anthropic chief product officer Mike Krieger, a recent hire and co-founder of Instagram.

Anthropic is no stranger to hyping up the prospects of AI.

In 2023, Dario Amodei predicted that so-called “artificial general intelligence” (capable of human-level thinking) would arrive within 2-3 years. At the end of 2024, he extended this horizon to 2026 or 2027.

He also estimated that AI will soon be writing most, if not all, computer code, making possible one-person tech startups with digital agents cranking out the software.

At Anthropic, already "something like over 70 percent of (suggested modifications in the code) are now Claude Code written", Krieger told journalists.

"In the long term, we're all going to have to contend with the idea that everything humans do is eventually going to be done by AI systems," Amodei added.

"This will happen."

GenAI fulfilling its potential could lead to strong economic growth and a “huge amount of inequality,” with it up to society how evenly wealth is distributed, Amodei reasoned.

K.Hashimoto--JT