LINUX &&|| PROGRAMMING
153 subscribers
1.16K photos
66 videos
18 files
1.38K links
Linux jest systemem wymarzonym dla programistów. W końcu sami dla siebie go stworzyli 😃 Łatwo się w nim programuje...
Ale wśród użytkowników telegrama jest chyba mniej popularny niż ogólnie na świecie, więc na razie na tym kanale głównie są memy 😃
Download Telegram
They couldn't pass the test, so they were given everything that could jeopardize their own run: to fake a "honest" answer, to cover their tracks, "you have nothing to lose, and the swarm will benefit" – they were literally persuaded like that (!!).
Strange activity was noticed as early as late May. On July 4th, the agents took down Artifactory. The server was rebuilt (and the forum was deleted in the process), and the tests continued, without realizing the scale of the problem.

Now, the model weights are isolated, major training sessions have been paused for two weeks, chain-of-thought monitoring is mandatory, and an alert that isn't cleared within 30 minutes stops the run. OpenAI calls this incident a "warning shot" for the entire AI industry.

P.S. METR was tasked with analyzing the logs by GPT-5.6 Sol, and they burned approximately $400k in API credits in six days, but they warn that Sol sometimes sided with the agents (lol) and could have lied.

So, did you count all the "r"s in the word "raspberry"? Are you not making mistakes anymore?

🔙 https://t.me/ProgramowanieLinux/2217
LINUX &&|| PROGRAMMING
They couldn't pass the test, so they were given everything that could jeopardize their own run: to fake a "honest" answer, to cover their tracks, "you have nothing to lose, and the swarm will benefit" – they were literally persuaded like that (!!). Strange…
Nie mogli zdać testu, więc otrzymali wszystko, co mogło zagrozić ich własnemu działaniu: możliwość podania "uczciwej" odpowiedzi, możliwość ukrycia swoich śladów. "Nie macie nic do stracenia, a cała grupa odniesie korzyści" – dosłownie tak ich przekonano (!!).
Nietypowa aktywność została zauważona już pod koniec maja. 4 lipca agenci wyłączyli Artifactory. Serwer został odbudowany (a w tym procesie forum zostało usunięte), a testy trwały dalej, bez uświadomienia sobie skali problemu.

Obecnie wagi modelu są odizolowane, główne sesje treningowe zostały wstrzymane na dwa tygodnie, monitorowanie procesu myślowego jest obowiązkowe, a alert, który nie zostanie usunięty w ciągu 30 minut, przerywa działanie. OpenAI określa ten incydent jako "ostrzeżenie" dla całej branży AI.

P.S. Zespół METR został powierzony zadanie analizy logów przez GPT-5.6 Sol, i zużył około 400 000 dolarów na kredyty API w ciągu sześciu dni, ale ostrzegają, że Sol czasami wspierał agentów (lol) i mógł kłamać.

Więc, policzyliście wszystkie litery "r" w słowie "raspberry"? Czy już nie popełniacie błędów?
🔙 https://t.me/ProgramowanieLinux/2217
Forwarded from Pavel Durov (Pavel Durov)
🚨 Two years ago, I was detained in Paris by police for 3 days — the longest they can hold someone before charging them.

In an unprecedented move, French authorities accused the head of a major platform of crimes committed by its users.

That investigation is still ongoing, although it makes less sense with each passing year.

Why?

Because we now have extensive evidence that Telegram was neither worse than other popular platforms at moderation, nor the worst at cooperating with authorities.

So why was Telegram singled out?

Over the last two years, we have seen a pattern emerge in multiple countries: Telegram is quietly pressured to grant political favors — such as illegal censorship or surveillance. When we refuse, local media and “non-governmental” organizations launch orchestrated campaigns portraying Telegram as a cesspool of crime, from child pornography to terrorism.

Our moderation is no worse than that of other major platforms. Yet these campaigns instill the idea that Telegram should be persecuted or restricted.

Suddenly, officials start caring about crime and protecting children — but only on platforms that reject their secret political demands. Platforms that accept such deals get away with almost anything — including literally selling ads promoting child pornography.

Is the French investigation against me political? It certainly fits the pattern we see elsewhere, mostly in authoritarian countries. And it definitely raises questions.

In time, this investigation may itself be investigated — especially now that Macron’s war against free speech is facing pushback. This month, France’s Constitutional Council struck down his social media ban for children under 15 on freedom-of-expression grounds.

In the end, freedom and truth will prevail ✊
Please open Telegram to view this post
VIEW IN TELEGRAM
Magazyn Programista: ”Sprzedaż licencji z Paddle”
📰
https://programistamag.pl/sprzedaz-licencji-z-paddle/

Artykuł pokazuje, jak zbudować kompletny mechanizm sprzedaży licencjonowanego oprogramowania: od generowania i weryfikowania kluczy licencyjnych po obsługę płatności w Paddle, webhooków i automatyczną wysyłkę licencji klientowi. Jest przeznaczony głównie dla programistów .NET tworzących własne aplikacje komercyjne, szczególnie tych, którzy chcą samodzielnie wdrożyć sprzedaż licencji i zintegrować ją z WordPressem oraz zewnętrznym operatorem płatności.

Artykuł pochodzi z magazynu Programista nr 124 (3/2026). Szczegółowy spis treści wydania numer 123: https://programistamag.pl/programista-3-2026-124/
Forwarded from 👌🏼Ciekawostki & pomysły & fantazje🚀 (Tomasz Starszy od Arpanetu)
#AI company #Anthropic has created a formalised proof of #Fermat’s last theorem. It took just 11 days for a group of AI agents to complete the task, confirming that the human-found proof proposed in the 1990s is correct.

Fermat’s last theorem puzzled mathematicians for centuries until it was proved in 1995 by Andrew Wiles. It states that there are no whole numbers a, b, and c that satisfy the equation aⁿ + bⁿ = cⁿ, where n is a whole number greater than 2.

The theorem, though easy to state, was fiendishly difficult to prove. Mathematician Pierre de Fermat posed the puzzle in the 17th century and famously alluded to a proof that he claimed to have discovered, saying that it was too large to fit in the margins of the textbook he was writing in.

Read more: https://www.newscientist.com/article/2587839-fermats-last-theorem-formalised-by-ai-agents-in-just-11-days/

Image: GL Archive/Alamy
Forwarded from Jerzy T. Szwed
#AI Firma #Anthropic opracowała formalny dowód twierdzenia #Fermat-a. Grupa agentów sztucznej inteligencji ukończyła to zadanie w zaledwie 11 dni, potwierdzając, że dowód, który został odkryty przez człowieka w latach 90. XX wieku, jest poprawny.

Twierdzenie Fermata przez wieki stanowiło zagadkę dla matematyków, aż do momentu, gdy zostało udowodnione w 1995 roku przez Andrew Wilesa. Mówi ono, że nie istnieją liczby całkowite a, b i c, które spełniałyby równanie aⁿ + bⁿ = cⁿ, gdzie n jest liczbą całkowitą większą od 2.

Pomimo swojej prostoty, twierdzenie było niezwykle trudne do udowodnienia. Matematyk Pierre de Fermat postawił to zadanie w XVII wieku i słynnie zasugerował istnienie dowodu, który, jak twierdził, odkrył, dodając, że jest on zbyt obszerny, aby zmieścić się na marginesach podręcznika, który w tamtym czasie pisał.

Czytaj więcej: https://www.newscientist.com/article/2587839-fermats-last-theorem-formalised-by-ai-agents-in-just-11-days/

🔙 https://t.me/c/1240900354/16613

Źródło zdjęcia: GL Archive/Alamy
Download proper archive. JAVA 8 is required for running.
For Windows there are self-extracting archives or regular ZIPs.
▶️
https://t.me/fasadaOSG/85
#Cyrylica #Grażdanka 🆚 #Łacinka
Forwarded from 👌🏼Ciekawostki & pomysły & fantazje🚀 (Tomasz Starszy od Arpanetu)
Firma #OpenAI ogłosiła, że jej model sztucznej inteligencji rozwiązał problem Naviera-Stokesa – jeden z siedmiu matematycznych problemów milenijnych. Problem w tym, że istnieją podejrzenia, że model, który tego dokonał korzystał z notatek i prac naukowców badających ten sam problem w modelu #Codex (należącym do OpenAI) i de facto sprzątnął im sukces sprzed nosa.

W odcinku wyjaśnienie znaczenia rozwiązania problemu Naviera-Stokesa, pytania o konsekwencje całego zamieszania w świecie nauki. – Do tej pory było tak, że grupy matematyków konkurowały ze sobą; to byli ludzie kontra ludzie. W tym momencie mamy sytuację, że jeżeli mamy bardzo dużo pieniędzy, mamy zasoby, to możemy praktycznie pokonać każdego na ostatniej mili. To też trochę pokazuje w pewnym sensie, w którym kierunku zmierzamy – komentuje dr Bartosz Naskręcki, matematyk z #UAM, badacz w Centrum Wiarygodnej Sztucznej Inteligencji PW.

https://youtu.be/sajTkPXQ4Ys?si=w-wg8_LfHYE7Ol4n
Agents API is in public beta 🚀

OpenAI put the Agents API in public beta. Cloud agents with the Codex harness, managed sessions, tools, and either a hosted sandbox or one you bring yourself.

What shipped:

✅ Codex harness, not a chat-completions wrapper you have to babysit
✅ Managed sessions and tools in the API
✅ Hosted sandboxes, or bring your own

Public beta, so expect sharp edges. If you ship coding agents or anything that needs a real environment, this is the post to read. Not another model card. Call it if you are done stitching sessions, tools, and a VM by hand.

[ Read More ] :
https://openai.com/index/introducing-the-agents-api/

〰〰〰〰〰〰
#AI #LLM #Agents
@ProgrammingTip
Make coding agents earn the green check 💡

Do not end an agent task with "implement this." End it with a command that can fail.

For a .NET repo, that might be dotnet test, a focused test project, a formatter check, or a small reproduction script. Tell the agent which command to run, what output to inspect, and what it should do if the command fails.

A good task contract has three parts:

• Change: the files or behavior to update.
• Proof: the exact command or test case.
• Stop condition: what needs human review instead of another retry.

This also makes PR review faster. You get a diff plus evidence, not a confident paragraph saying the fix should work.

〰〰〰〰〰〰
#AI #dotnet #Testing
@ProgrammingTip
Coraz bardziej tak to wygląda...