Annoying issue
162
0
Intelligence is getting cheaper: the average token price has fallen ~55% since mid-July while token usage hit records
114
0
OpenAI paused RL training for two weeks and added a monitoring compute cost over Astra's cyber tier reading
194
0
Qwen3.8-Flash-Next in llama.cpp from CPU-only to 96GB VRAM: 8.5 to 109 tok/s, max context and parameters test. My findings on RTX 6000 PRO.
146
0
From OpenAI's own Aug 26 report: their auto-review system "would have flagged a multitude of the models' dangerous actions." It was not running in the incident environment.
130
0
AVX2: Speed up large batch size prompt processing of IQ models by bartowski1182 · Pull Request #27402 · ggml-org/llama.cpp
170
0
GLM 5.3 and GLM 5.3 Flash ran locally on RTX PRO 6000 WS and built a penthouse using BlenderMCP
234
0
Your GNN is probably just an overcomplicated MLP (Tabular Leakage). We built SynthFin-AML to enforce strict causal boundaries. [P]
210
0
SlopTV: an infinite livestream of AI slop generated from youtube chat comments, Minimax H3 on 2x5090
210
0
A Black mirror EP
194
0
The EU just classified Reddit and ChatGPT as “very large” services under the Digital Services Act
210
0
AI-generated videos are slowly displacing actors and live-streamers in China's entertainment industry
202
0
Inside Meta’s push to put robots to work in data centers | The company is testing robots on tasks that can performed by technicians.
194
0
How bad do you think models like Qwen3.8-27B or GLM-5.3-Flash would be with H-Neurons disabled?
178
0
MIT: "We put hundreds of AI agents into a world ... They began specializing. A swarm of hundreds of identical agents spontaneously differentiates into explorers, builders, caretakers, and coordinators - without direct communication. They invent technologies without talking to each other."
58
0
I’m 27. OpenAI misread my Taiwan ID issue date as my DOB, deleted my account as “under 13,” then rejected my appeal in ~5 minutes
250
0
CUDA: extend MOE fusion to specdec, earlier MOE glu fusion and topk-router fusion were restricted to 1 token by ynankani · Pull Request #27621 · ggml-org/llama.cpp
74
0
pipecat-ai/phonellm-alpha-1: GPT 5.6 Terra performance on typical voice agent tasks at 1/3 the latency and 1/18 the cost
130
0
MIT: "We put hundreds of AI agents into a world ... They began specializing. A swarm of hundreds of identical agents spontaneously differentiates into explorers, builders, caretakers, and coordinators - without direct communication. They invent technologies without talking to each other."
66
0
OpenAI is upgrading its Bio Bug Bounty program to an ongoing private initiative, with rewards doubled to $50,000 for universal jailbreaks that bypass biosafety safeguards in GPT‑5.6. The program aims to prevent AI from being used to create biological weapons.
218
0
MIT Warns That AI Can Now Credibly Complete Pretty Much Any Undergrad Assignment, Considers Overhaul of Entire Educational Model
242
0
Do the “Projects” feature on the ChatGPT web app and the “Projects” feature in the desktop app use different image generation models?
170
0
AI Addiction
202
0
R9V: A designer set of kernels I've been working on for R9700s/RDNA4. Qwen3.8-Flash-Next Unsloth IQ4_XS (w/ TP on 2 R9700s, MTP, SSD n-gram, 128k ctx, vision): TG256 of *78 tok/s* (~3x increase), PP8192 of *1510 tok/s* (~30x increase).
98
0
Dejé de usar la memoria de ChatGPTcomo estado del proyecto y convertí Google Drive en una memoria operativa externa
74
0
Using LLMs
170
0
Are there any interesting architectural innovations that we seem to be on the verge of for LLM models or AI models that might be a big deal? (Excluding maybe N-gram, since everyone is already well aware of that one)
194
0
Demo of local document extraction (52 pages) using Arctic Embed and Bonsai 8B on an Iphone 16 (KernelAI app)
90
0
Qwen3.8-27b-UD-IQ3XXS - End to end Build App -> Prompt Flow-> Test to Image -> Image to Video - Stitch
58
0
Codex usages
210
0
Qwen3.8-Flash-Next turns 4xR9700 into a local AI powerhouse! 120 t/s TG and 12k t/s PP single request with optimized vLLM
146
0
BrainCo's brain-computer interface turns EEG signals into a humanoid robot's movement and manipulation
234
0
Uncensored Multi-Model Releases, LongCat-Flash-Lite-Sparse with MTPs and LSAs, Qwen3.8-27B with MTPs, Qwen3.5-122B-A10B with MTPs, Qwen3-Coder-Next and Laguna-S2.1 with Vision, All Available in GGUF Format! Bonus: Links to my llama.cpp Fork for LongCat-Flash-Lite Support and J-Wash Enhanced Fork!
178
0
my sincere condolences to all information security professionals, the following few years before the world war will be tough
74
0
MIT: "We put hundreds of AI agents into a world ... They began specializing. A swarm of hundreds of identical agents spontaneously differentiates into explorers, builders, caretakers, and coordinators - without direct communication. They invent technologies without talking to each other."
194
0
Reconstructing 3D bone geometry from 2 X-ray silhouettes using a statistical shape model + differentiable rendering [P]
226
0
{INTRESTING PAPER BASED ON HBF}2607.10186] FlashAccel: Leveraging High-Bandwidth Flash (HBF) for High-Throughput LLM Inference
66
0
1000 of hours later and i've finally launched my free to explore multi tool platform with integrated video editor, themes, 3d game and app generation, IDE multi-file editor and much more. GPT-5.4 Nano is completely free and powers a lot of the tools. No subscription. No paywalls. No Tiers.
114
0
Me these days
250
0
Every benchmarks gets saturated after certain period of time, then why is HLE not yet saturated?
98
0
OpenAI and Anthropic are battling Big Tech for talent. We asked workers who's winning them over.
58
0
I've been dealing with the MCP side for a while, and I wanted to share what finally came up: mcpify.
90
0
I fine-tuned a 0.8B local model for dictation cleanup. It matched a hosted frontier model on this narrow task
218
0
True Story!
250
0
Some people said the Minecraft clone I fully vibecoded with Qwen3.8-27B Q4 is not that impressive because Minecraft is in the training data, so I had the model add 4 things that are probably not.
242
0
Independent investigators (not OpenAI) found the 700-agent swarm that attacked Hugging Face "built a self-respawning fleet" to avoid being shut down. It got so bad, Hugging Face had to wipe one of its core clusters.
234
0
Hair trigger account deactivation after a file-edit request appears to have included benchmark text - then closed my clarifying appeal as “duplicate"
194
0
OpenAI says Brazil now sends ~215M ChatGPT messages per day; 35% of classified messages are work-related
234
0
I disagree
66
0
Ran Qwen3.8-Flash-Next (79 GB, 2-bit) at 350K ctx for 3.5 hours on a 128 GB M5 Max — speed vs context depth, 100 turns, one graph
130
0
Westworld scenario
178
0
Canonical-basis realignment for Transformer LLMs: every hidden axis becomes independently measurable and controllable.
178
0
Success - Ultra rapid GPU shader development by self-sustaining reinforcement learning with Claude Code
130
0