Lotu RadarAbout · RSS

Latest News Archive - Page 111

World · The Guardian Ukraine

Ukraine war briefing: Russian sky full of our drones and not safe, Zelenskyy warns airlines

Germany accuses Russia over Leipzig airport attack, citing solid evidence; Russian economy in risky ‘berserk mode’ says Putin development tsar. What we know on day 1,652 Continue reading...

World · Deutsche Welle

Ukraine: Russia threatens 'devastating' Baltic response

Any escalation in the Baltic region will lead to a significant response, Moscow said. Follow DW for more.

Developers & Open Source · ollama/ollama Releases

v0.33.3

What's Changed Honor GGUF model defined default parameters MLX, MLX-C, llama.cpp update New Contributors @marcelpetrick made their first contribution in #17579 Full Changelog : v0.33.2...v0.33.3-rc0

World · Al Jazeera

Iran war live: US bombs Iran, Tehran retaliates on Gulf neighbours, Jordan

Tehran pledges 'severe punishment' and retaliation against Washington following a new wave of US attacks.

Notable Blogs · Simon Willison

Claude Fable 5.1 made me a really nice animated pelican

Today is Claude Fable (and Mythos) 5.1 day . Anthropic say that Fable 5.1 "sets a new standard for coding, knowledge work, and long-running problem-solving tasks". Their announcement spends a notable amount of time on scientific research, boasting of a 52.6% score on the brand new Terminal-Bench-Science 0.1 benchmark (first announced on August 27th ), up from 24.7% for Fable 5, 29.0% for Opus 5 and 22.4% for GPT-5.6 Sol. Other benchmarks show slightly improved scores, but none as impressive as the Science one. But how well can it pelican? Back in July I wrote about how I was losing faith in the pelican benchmark - its connection to how good the models were at other tasks didn't seem to hold as strongly as it did back in 2025 . The most interesting insights I get from it now are comparisons within model families, and particularly comparisons for the same prompt at different reasoning effort levels. Fable 5.1 has five reasoning levels: low, medium, high, xhigh, max - and no option to turn off reasoning entirely. I fixed an issue in llm-anthropic which caused reasoning traces not to be correctly recorded, then ran some prompts. Here's the full set of pelicans for all of the reasoning levels, each with the full reasoning transcript. I'll replicate them here: Low and medium, both without reasoning? Next, a bit of a mystery. This is what I got for effort low : The transcript doesn't show any summarized reasoning tokens, and the output token count is 1,998. With Claude that output token count includes reasoning tokens. It took 23.8 seconds and cost 10.017 cents . I bumped that up to medium and got this: Weirdly, that one also shows no reasoning text and used 1,977 output tokens - 21 tokens less than low . It took 23 seconds and cost 9.912 cents . So for this particular prompt ("Generate an SVG of a pelican riding a bicycle") Fable 5.1 appeared to skip reasoning entirely at both low and medium settings. High Here's high - 29.6 seconds, 2,612 output tokens, 13.087 cents : This one did do a bit of reasoning, summary here : I'm planning the SVG layout for a pelican riding a bicycle, with a sky and ground background, a bicycle with two spoked wheels, frame, seat and handlebars, and a white-bodied pelican with a long neck and orange beak positioned on top. Really not much difference from low and medium , though. Extra High At xhigh things got radically different. 36,767 output tokens, 7 minutes 51 seconds, $1.83 ! The reasoning trace is pretty lengthy , and includes details like this: Adding the eye, wings stretching down to the handlebar grip, orange legs reaching to the pedals, and a small tail feather, while keeping the pelican intentionally oversized compared to the bike for comic effect. [...] I'll accept the slight thickness as charming rather than overengineering it. Max Setting effort to max gave me the best pelican I've seen from any of Anthropic's models. 65,927 output tokens, 13 minutes and 54 seconds, $3.30 : There's a lot to like about this. The background is tasteful, the legs are clearly on either side of the frame, the feet are on the pedals, the wing is on the handlebars, the pelican has a cute blue hat and there's a basket with a fish. It's still not showing nearly the same level of flair as Gemini 3.7 Flash , but I didn't ask for flair - I asked for an SVG, and that's what I got. Some highlights from that reasoning trace : Adding pedal shapes near both feet, with the far foot on the second leg partially visible behind the frame. I'm considering whether to add a small scarf or cap for extra character, but leaning toward keeping it simple to avoid clutter. Now I'm debating a bicycle helmet on the head versus the pelican's signature crest—the beak and pouch already read clearly as "pelican," so a helmet could reinforce the bicycle theme without losing identity, though it might compete with the crest for visual space. I realize the beak at (484,84) would overlap with the dome helmet, so I need to shrink the helmet so it only covers the top of the head, adjusting its arc endpoints to sit higher and narrower so the beak can attach cleanly at the front without collision. [...] I'm adding a darker tip region to represent the primary feathers, then reconsidering the trailing edge to include scalloped feather curves instead of one smooth line for a more natural look. [...] Now I'm checking the vent line placements on the helmet, making sure they sit far enough inside the helmet's edge given the stroke width and rounded caps, and confirming each vent stays within the helmet's circular boundary. [...] I decide skipping a handlebar bell and tire highlights since they're unnecessary additions. Now I'm reconsidering the front fork's curve — the current control point pulls the shape backward when it should bow forward for a proper rake, so I need to shift the control point rightward to fix the fork's lean. OK, let's animate it On Hacker News, swalsh commented on that Max pelican: Now that it's a solved benchmark, can we get the animated version? I didn't want to spend another $3 so I took the Max pelican and piped it into the default thinking level of High: llm logs -cx | llm -m claude-fable-5.1 -s ' animate this ' 6,121 input, 26,201 output = $1.37 . The result looked like this , exported here as video since some people have trouble viewing animated SVGs: Your browser does not support HTML5 video. The wheels in the video are rotating in the wrong direction, but I think that's an artifact of the conversion to MP4 - they seem to be going in the correct direction in the original SVG. Tags: ai , generative-ai , llms , anthropic , claude , pelican-riding-a-bicycle , llm-reasoning , llm-release

World · BBC Middle East

Iran says US strike killed four at wedding in 'war crime' as US denies targeting civilians

Iranian media reports two children were among those killed when shrapnel hit a home. Iran then launched missiles and drones at US targets in the Middle East.

Developers & Open Source · vercel/next.js Releases

v16.4.0-canary.14

Misc Changes turbo-tasks-malloc: report memory from mimalloc: #97761 Remove Edge runtime handling in use-cache-wrapper: #98139 Skip discussions in the issue_lock workflow: #98146 [ci] Fix issue_lock : #97755 docs: document preloading with Cache Components: #97864 chore: Remove expired sharp release age exclusions: #98120 Credits Huge thanks to @lukesandberg , @mischnic , @eps1lon , @aurorascharff , and @styfle for helping!

Society · BBC Technology

Will self-flying planes transform the skies?

Autonomous crop-spraying aircraft are leading the way in pilot-free flying.

Business · BBC Business

Soft launches and late sittings - six ways to get cheaper meals out

The cost of dining out is getting increasingly hard to swallow - but clever planning can net you cheaper deals.

Business · CNBC Technology

Palo Alto CEO says $1 trillion of cybersecurity infrastructure isn’t ready for AI

Palo Alto CEO Nikesh Arora said AI is forcing companies to modernize roughly $1 trillion of aging cybersecurity infrastructure that isn’t equipped for attacks.

Products & Consumer Tech · Ars Technica

Here's our first look—and drive—of the 2027 Range Rover Electric

333 miles of real-world range with uncompromised comfort and off-road ability.

World · Al Jazeera

Fernandez transfers to Man City from Chelsea in joint British record fee

Argentina international Enzo Fernandez signs for Manchester City in a deal from Chelsea worth £125 million ($169m).

Developers & Open Source · Chrome Releases

Chrome for Android

 Hi, everyone! We've just released Chrome 152 (152.0.7977.75) for Android. It'll become available on Google Play over the next few days.  This release includes stability and performance improvements. You can see a full list of the changes in the Git log . If you find a new issue, please let us know by filing a bug . Android releases contain the same security fixes as their corresponding Desktop releases (Windows & Mac: 152.0.7977.75/76 Linux: 152.0.7977.75) unless otherwise noted. Krishna Govind Google Chrome

Products & Consumer Tech · The Verge

Google needs Hollywood more than the studios need AI

Google has reportedly been reaching out to a number of Hollywood's biggest studios, hoping to strike licensing agreements that would allow it to train its AI models on copyrighted material in exchange for massive piles of cash. In theory, these deals would be a win-win: a huge financial boon to the studios that would also […]

Cybersecurity · Krebs on Security

FBI Probes Service Selling 153M+ Drivers Licenses

A new identity theft service launched on the dark web this week is selling digital scans of more than 153 million drivers licenses from people in the United States and Canada. Based on interviews with individuals whose licenses are available for purchase on this service, it appears to be siphoning images collected by a widely-used identity verification company based in Louisiana. KrebsOnSecurity also has learned that the New Orleans field office of the Federal Bureau of Investigation (FBI) today launched an official inquiry into the source of the images.

Developers & Open Source · GitHub Changelog

Enterprise Live Migrations from GHES to ghe.com generally available

Enterprise Live Migrations (ELM) is now generally available, enabling near-zero-downtime repository migrations from GitHub Enterprise Server (GHES) to GitHub Enterprise Cloud with Data Residency (GHEC DR), so you can move… The post Enterprise Live Migrations from GHES to ghe.com generally available appeared first on The GitHub Blog .

Society · NPR Top Stories

North Carolina Rep. Chuck Edwards is formally censured over harassment allegations

The censure vote drew bipartisan support, and followed an investigation by the House Ethics Committee into allegations of sexual harassment against two young female staff members.

Developers & Open Source · Chrome Releases

Beta Channel Update for ChromeOS / ChromeOS Flex

The Beta channel is being updated to OS version 16765.38.0 (Browser version 152.0.7977.74 ) for most ChromeOS devices. If you find new issues, please let us know one of the following ways: File a bug Visit our ChromeOS communities General: Chromebook Help Community Beta Specific: ChromeOS Beta Help Community Report an issue or send feedback on Chrome Interested in switching channels? Find out how. Luis Menezes Google ChromeOS  

AI · TechCrunch AI

AfterQuery reportedly becomes Y Combinator’s fastest-ever unicorn, now valued at $3.2B

AI model-training startup AfterQuery has reportedly raised a round that valued it at $3.2 billion, just five months after announcing its $30 million Series A at a $300 million valuation in April.

Products & Consumer Tech · The Verge

Anthropic launches Claude Fable 5.1 and says it’s up to 45 percent cheaper for agentic work

Anthropic says its newest AI models, Fable 5.1 and Mythos 5.1, address criticisms from customers about price, data retention, and overzealous safeguards. The company claims Claude Fable 5.1 offers stronger performance than Fable 5, but costs around 25 percent less typically and up to 45 percent less for complex agentic tasks, thanks to reduced pricing […]

Business · CNBC Technology

Palo Alto Networks beats quarterly estimates on AI demand, continues acquisition spree

Palo Alto Networks' stock has nearly doubled this year as AI boosts demand for security detection and response tools

Developers & Open Source · Hugging Face Blog

BenchMIRT: What are LLM benchmarks actually measuring?

Hugging Face Blog published: BenchMIRT: What are LLM benchmarks actually measuring?

Products & Consumer Tech · Product Hunt

FreeScan.app

Fix what’s hurting your visibility, trust, and conversions Discussion | Link

Products & Consumer Tech · Product Hunt

Relaticle

Open-source CRM with approval-gated AI writes Discussion | Link