🏷 Tag

llm · 1206 topics

papers (249)

2026 · Sep

2026 · Aug

08-31🔥🔥

evolution strategies for llm reasoning broader coverage

08-31🔥🔥

the mask is not the model

08-31🔥🔥

best practice critic optimization

08-30🔥🔥

vgi bench video generation visual intelligence

08-30🔥🔥

jit agent harness evolution

08-30🔥🔥

frontierchallenge evaluating scientific workflow completion

08-30🔥🔥

wemm embedding

08-29🔥

eduriskx neuro symbolic framework

08-29🔥

icu mortality llm agentic pipeline

08-29🔥

large models for battery prognostics and health management

08-28🔥🔥

criticl weak to strong generalization

08-28🔥🔥

wikiskill compiling agent experience into persistent knowledge

08-28🔥🔥

ttpo test time policy optimization

08-28🔥🔥

swe prime fewer trajectories better performance

08-28🔥🔥

llms in hpc programming survey

08-28🔥🔥

treegraft adaptive multi drafter grafting

08-28🔥🔥

label free doubt signals for abstention

08-28🔥🔥

agents dont paginate first chunk selection for llm tool responses

08-27🔥🔥

vbvr pro native visual reasoning

08-27🔥🔥

agentic autoresearch for cell edge power control

08-27🔥🔥

prefix sliding

08-27🔥🔥

how much rank does lora need rank error bounds for transformer attention

08-26🔥🔥

triplu trilinear product ffns tiny language models

08-26🔥🔥

multilingual verifier bias in rlvr

08-26🔥🔥

va dpo valence arousal direct preference optimization for controllable emotion generation in language models

08-25🔥🔥

primeagentorchestrator memory primed agent spawning for personal ai infrastructure

08-25🔥🔥

mechanistic analysis of occupational bias

08-25🔥🔥

inhibitory attention clinical long context reasoning

08-25🔥🔥

beyond prompt engineering lexical sensitivity

08-25🔥🔥

clarify then search benchmark

08-24🔥🔥

semaplc plc code generation

08-24🔥🔥

semcomp bench semantic task completion video generation

08-24🔥🔥

spade self play adaptive synthetic executable environments

08-24🔥🔥

training chemical plausibility aware large language models for single step retrosynthesis

08-23🔥🔥🔥

moss vl technical report

08-23🔥🔥

swe bench science

08-23🔥🔥

co rl unsupervised reasoning multi agent rl

08-23🔥🔥

agentic esopt fine tuning long horizon llm agents with minimal gpu requirements

08-23🔥🔥

saturation aware advantage reweighting

08-22🔥🔥

fm bench

08-22🔥

transformer models text summarization

08-22🔥

automatic bioinformatic software named entity recognition

08-21🔥🔥

sub billion entity tracking

08-21🔥🔥

sutra structurally unified tokenization

08-21🔥🔥

latent space refusal anchoring

08-21🔥🔥

nine emotion centroids label free valence axis

08-21🔥🔥

fractional decay kv cache

08-20🔥🔥

intent driven dynamic chunking

08-20🔥🔥

gxp agent process dag topology

08-20🔥🔥

when personalization becomes bias

08-20🔥🔥

hierarchical data selection mass

08-20🔥🔥

constant competitive algorithm for dynamic mixture of experts serving

08-19🔥🔥

trace temporal reasoning benchmark

08-19🔥🔥

auxiliary uncertainty signals for llm assisted systematic review screening

08-19🔥🔥

forward pass domain adaptation

08-19🔥🔥

ogx open source vendor neutral generative ai application server

08-18🔥🔥🔥

from bert to frontier agents

08-18🔥🔥

proxy validated llm ux micro simulations

08-18🔥🔥

rubricforge agent evaluation

08-18🔥🔥

depth aware sensitivity analysis of moe models

08-18🔥🔥

modular cognitive architecture emerges in large language models

08-17🔥🔥

rhetorical sensitivity ai peer review

08-17🔥🔥

u opsd unsupervised on policy self distillation

08-16🔥🔥

spark to paper

08-16🔥🔥

ai4ai test time capability transfer

08-16🔥🔥

mechanist ai scientific instrument mechanisms intelligence

08-16🔥🔥

bdh cq in context learning with recurrent latent reasoning

08-15🔥🔥

lora diffusion parameter efficient fine tuning via low rank trajectory decomposition

08-15🔥🔥

thought aware kv cache compaction for reasoning via adaptive attention matching

08-15🔥🔥

financial error detection benchmark fined bench

08-15🔥🔥

dual flow transformers

08-14🔥🔥

evaluating llm generated detection rules in cybersecurity

08-14🔥🔥

wavephasenet dft schs

08-14🔥🔥

llm reasoning reliability reproduction

08-14🔥🔥

dynamic governance of multi llm agent systems

08-14🔥🔥

distribird literature informed prior distribution design

08-13🔥🔥

divergent response modes in frontier language models

08-13🔥🔥

llm agents factory retrieval of domain specific llm agents

08-13🔥🔥

how to dogfood your ai chat agent

08-13🔥🔥

multilingual quantization tax

08-13🔥🔥

llm cot serial depth bottleneck

08-12🔥🔥

from trajectories to evidence auditable experimental records for industrial research agents

08-12🔥🔥

llm agents social media reaction prediction

08-12🔥🔥

the judge knows when it knows

08-12🔥🔥

training variable long sequences with data centric parallel

08-11🔥🔥

sharding prevents llm oversight failures

08-11🔥🔥

cyberforge repository level vulnerability injection

08-11🔥🔥

measuring the cross lingual comprehension gap

08-10🔥🔥

joyai video edit

08-10🔥🔥

economic world models systems blueprint

08-10🔥🔥

the personalization mirage

08-10🔥🔥

video deepresearch

08-10🔥🔥

hunyuan3d buffalo 1 0

08-09🔥🔥

recursive synthetic terminal tasks

08-09🔥🔥

abseeker long horizon search agents

08-09🔥🔥

toolartist

08-09🔥🔥

towards physics of multimodal pretraining

08-08🔥🔥

nvidia b300 multi node full fine tuning

08-08🔥

llms preferences for libraries and programming languages

08-08🔥

flowadam geometry aware soft momentum injection

08-08🔥

guiding llms with gp evolved knowledge for project scheduling

08-07🔥🔥

timestep conditioned transformers for global weather forecasting

08-07🔥🔥

what current ai benchmarks leave unmeasured

08-07🔥🔥

llm conspiracy reduction

08-06🔥🔥

agentstream self evolving llm agents streaming

08-06🔥🔥

kernelbrain coarse to fine budget aware search for agentic gpu kernel optimization

08-06🔥🔥

memarena ego centric benchmark on device agentic memory

08-06🔥🔥

evaluating openais privacy filter cross lingual cross domain pii detection

08-04🔥🔥

evidence ledger adjudication

08-04🔥🔥

svr self verifying refinement

08-04🔥🔥

openclaw ollama agentic ai

08-04🔥🔥

ai scientist evaluation benchmark

08-04🔥🔥

llm framework for discovering major mathematical conjectures

08-03🔥🔥

bm25 wins at scale rag scaling study

08-03🔥🔥

flux opd on policy distillation

08-03🔥🔥

videococo code as cot

08-03🔥🔥

codenib multi view data system

08-03🔥🔥

cort counterfactual replay for token level rubric guided policy optimization

08-02🔥🔥

qwen ui agent technical report

08-02🔥🔥

frontis ma1 recursive self improvement mle

08-02🔥🔥

metis memory foundation model

08-02🔥🔥

turbovla real time vla model

08-02🔥🔥

memory decoder at scale

08-01🔥🔥

magicselector joint optimization for agent tool selection

08-01🔥🔥

guidedrag semantic steering of retrieval augmented generation

08-01🔥🔥

gpt red automated red teaming self play

08-01🔥🔥

probing origins of reasoning performance

08-01🔥🔥

kernelgenbench multi source multi chip benchmark

2026 · Jul

07-31🔥🔥

docannot lvlm kie dataset generation

07-31🔥🔥

do models fake alignment without clear consequences

07-31🔥🔥

kernel forge agent harness cuda

07-31🔥🔥

llm scheming inversely scales with pretraining language coverage

07-30🔥🔥

codifying the judge

07-30🔥🔥

craft learn the schema execute the plan

07-30🔥🔥

beyond shapley influence based data auditing pipeline for llm alignment and evaluation

07-30🔥🔥

fusionml cpu gpu co execution apple silicon

07-29🔥🔥

evaluating large language models for symbolic security protocol analysis

07-29🔥🔥

semalith v1 4 safety classifier

07-29🔥🔥

deeplens diagnosis agent

07-29🔥🔥

sf ams structured memory llm agent

07-29🔥🔥

medlocomo long context medical dialogue benchmark

07-28🔥🔥

arex recursively self improving agent deep research

07-28🔥🔥

rubric oriented document set selection and ranking

07-28🔥🔥

tencent workbuddy bench

07-28🔥🔥

nvidia labs oo agents

07-27🔥🔥🔥

slai t rex deepseek v4 ascend superpod

07-27🔥🔥🔥

rynnbrain 1 1

07-27🔥🔥

mage flow native resolution foundation model

07-27🔥🔥

swe pruner pro

07-26🔥🔥

streaming multi agent autoregressive diffusion model with world state registers

07-26🔥🔥

openforgerl harness native agents

07-26🔥🔥

mirror learning from the other view for multi modal reasoning

07-25🔥🔥

is moe routing a huffman code

07-25🔥🔥

human in the loop llm cirAE identification

07-25🔥🔥

skill contracted agents for evidence aware materials literature analysis

07-25🔥🔥

topoguard graph theory based defenses against split knowledge attacks on rag

07-25🔥🔥

glan qna kr seedless taxonomy driven korean instruction corpus

07-24🔥🔥

basert apple m5 inference

07-24🔥🔥

slpo scaling latent reasoning

07-24🔥🔥

triagent divergence aware multi agent committees

07-24🔥🔥

evothink evolving thinking in large reasoning models

07-24🔥🔥

solar open 2 technical report

07-23🔥🔥

sysadmin measuring instrumental power seeking in frontier ai

07-23🔥🔥

a controlled study of attention only transformers

07-23🔥🔥

relay bench evaluating llms on multi domain reasoning chains

07-23🔥🔥

structured output collapses answer diversity

07-23🔥🔥

search on graph r1

07-22🔥🔥

rater state bias in rlhf preference data

07-22🔥🔥

planflip attacking multi agent llm systems

07-22🔥🔥

deterministic replay for ai agent systems

07-22🔥🔥

rail guard closing the evaluation to remediation gap

07-21🔥🔥🔥

hidden in thought

07-21🔥🔥

beyond a single direction

07-21🔥🔥

cura 1t specialized model for agentic healthcare

07-21🔥🔥

xiaomi robotics 1 scaling vla models

07-20🔥🔥🔥

unicode tag block concealment mcp

07-20🔥🔥

mcpevol bench benchmarking llm agent performance across dynamic evolutions of mcp servers

07-20🔥🔥

retroagent harnessing llms to search over structured memory

07-20🔥🔥

beyond generalist llms specialist agentic systems

07-20🔥🔥

catalogagent supervisor mediated self learning

07-19🔥🔥

toolanchor anchoring counterfactual context

07-19🔥🔥

reward free evolving agents via pairwise validator

07-19🔥🔥

muse representation geometry of muon beyond normalized momentum

07-18🔥🔥

hg rag hierarchy guided retrieval augmented generation for structured knowledge graphs

07-18🔥🔥

token time continuous diffusion for language modeling

07-18🔥🔥

polestar drift aware cache calibration

07-18🔥🔥

eta given delta

07-18🔥🔥

information theoretic limits of reliability and scaling in language models

07-17🔥🔥

citybehavex urban simulation

07-16🔥🔥

braille accessibility failures in llms

07-16🔥🔥

semidirect fourier delta attention

07-16🔥🔥

transforming llms into efficient cross encoders via knowledge distillation for rag reranking

07-16🔥🔥

litetopk exploiting the curse of dimensionality

07-14🔥🔥

contrastive learning on multimodal analysis of electronic health records

07-12🔥🔥

deepsearch world

07-10🔥🔥

agentlens

07-09🔥🔥

dynamics fine tuning llms

07-09🔥🔥

mt editflow

07-08🔥🔥

gemma 4 technical report

07-07🔥🔥

understanding annotator safety policy with interpretability

07-07🔥🔥

revisiting asr error correction with specialized models

07-07🔥🔥

path constrained mixture of experts

07-03🔥🔥

basert best in class llm inference on apple silicon via native metal

07-01🔥

memora harmonic memory

2026 · Jun

2026 · May

2026 · Apr

product (74)

2026 · Sep

2026 · Aug

2026 · Jul

2026 · Jun

2026 · May

2026 · Apr

research (494)

2026 · Sep

09-07🔥🔥🔥

gpt 6 astra on robot arms

09-07🔥🔥

k2 horizon mova 36b a4b

09-07🔥🔥

vdn minimax h3

09-07🔥🔥

nyu mll glue benchmark

09-07🔥

facebook mms 300m

09-07🔥

openai community gpt2

09-07🔥

claude system prompt song lyrics

09-06🔥🔥🔥

nvidia nemotron ioi 2026

09-06🔥🔥

davidau qwen38 27b turbo fable cold fusion

09-06🔥🔥

qwen3 8 27b gsq rco gguf

09-06🔥🔥

spark x25 4b release

09-06🔥🔥

artificial analysis intelligence index v4 2

09-06🔥🔥

cohere labs ate dataset reveals only 2 6 of ai agent tools actually work

09-05🔥🔥🔥

openai gpt 6 astra

09-05🔥🔥🔥

anthropic claude fermat

09-05🔥🔥🔥

openais rogue agents were caught communicating via public wikis

09-05🔥🔥

google opens lyria 3 5 to developers

09-05🔥🔥

google timesfm 3.0 pytorch pytorch

09-05🔥

faunix qwen38 27b distillation 40k

09-04🔥🔥🔥

openai gpt 6 astra computer use

09-04🔥🔥

xiaohongshu self gc context manager

09-04🔥🔥

mbzuai k2 horizon 375b

09-04🔥🔥

neomme multimodal encoder

09-03🔥🔥🔥

gemini 3 8 flash and 3 8 flash cyber

09-03🔥🔥🔥

claude fable mythos 51 launch

09-03🔥🔥

metas muse spark 1 3 cuts tool calls

09-03🔥🔥

google gemini 3 8 flash

09-03🔥🔥

multiverse computing quasar 438b

09-03🔥🔥

alibaba qwen3.8 max 0902 tops code arena

09-02🔥🔥🔥

claude fable 5 1 and claude mythos 5 1

09-02🔥🔥🔥

openai astra cybersecurity

09-02🔥🔥

benchmirt llm benchmarks

09-02🔥🔥

google deepmind gemini 3 7 flash video analysis

09-02🔥🔥

anthropic claude fable 5 1

09-02🔥🔥

metas muse voice transcribe beats every rival with 3 1 error rate

09-01🔥🔥

how to build a diffusion language model

09-01🔥🔥

anthropic trained hacker opus on hackable environments

09-01🔥🔥

runways solaris renders interactive apps as live video no code needed

09-01🔥🔥

nvidia claude science bionemo

09-01🔥🔥

anthropic ai safety research

09-01🔥🔥

apodex 1 1 beats deepseek on real world agentic work with a 35b open model

2026 · Aug

08-31🔥🔥

hallucination by proxy llm assisted diagnosis

08-31🔥🔥

unbounded self improvement and its limits

08-31🔥

fastvideo fasth3 4 step preview

08-31🔥

qwen3 8 flash next fp8

08-31🔥

specific labs scaffold cot

08-31🔥

pipecat ai phonellm alpha 1

08-30🔥🔥

minimax h3 fun controlnet union

08-30🔥🔥

anthropic enabling independent research

08-30🔥🔥

perplexity adds glm 5 3

08-30🔥🔥

krea reveals krea 3 and agents

08-30🔥🔥

autonomous mathematical discovery station

08-30🔥🔥

sapiens ai agnes 2 5 pro beta

08-30🔥🔥

llms are not consistently bayesian

08-29🔥🔥🔥

z ai glm 5 3 open weights

08-29🔥🔥

anthropic automated alignment researcher

08-29🔥🔥

epochs ebr bench reveals humans learn while frontier ai stays stuck

08-29🔥🔥

anthropic aar claude opus alignment

08-29🔥🔥

sesames turnbench exposes how gemini live and openai realtime

08-29🔥🔥

terminal bench science evaluating ai agents

08-29🔥🔥

tencent hy4 preview

08-29🔥🔥

agent seer synthesizing scenarios

08-28🔥🔥

google gemini omni 1 1 flash

08-28🔥🔥

cohere parse 5

08-28🔥🔥

google deepmind gemini flash lite double blind evaluation

08-28🔥🔥

glm 5 3 flash

08-28🔥🔥

glm 5 3 flash performance analysis

08-28🔥🔥

anthropic model hardware standard lab robots

08-28🔥🔥

sensenova u1.5 8b mot

08-28🔥🔥

unsloth glm 5 3 flash gguf

08-28🔥🔥

from preferences to principles

08-28🔥🔥

goodfires forking fast slashes ai reasoning debug costs by 100x

08-28🔥🔥

perplexitys brain beats vector search with a self rewriting memory wiki

08-27🔥🔥🔥

qwen38 flash next

08-27🔥🔥

z ai glm 5 3 flash matches claude opus

08-27🔥🔥

z ai ox alpha model

08-27🔥🔥

googles glucofm predicts diabetes

08-27🔥🔥

glucofm foundation model for continuous glucose monitoring

08-27🔥🔥

anthropic insights real claude conversations study

08-26🔥🔥

agenthands generating interactive hand gestures for spatially grounded agent conversations in xr

08-26🔥🔥

aletheia quest retrospective

08-26🔥🔥

granite 4 2 llms

08-26🔥🔥

quantization aware healing

08-26🔥🔥

mozilla state of open source ai

08-26🔥🔥

starflow2 bridging language models and normalizing flows

08-26🔥🔥

liquid ai pipette on device benchmarking

08-25🔥🔥🔥

orcarouter qwen3 8 27b uncensored gguf

08-25🔥🔥

ai agent skills study

08-25🔥🔥

liquid ai lfm2 5 on device phone benchmarks

08-25🔥🔥

30b open models trend

08-25🔥🔥

apple internalized visual thinking ivt

08-25🔥🔥

qwen3 8 max glm5 2 kimi k3 distillation

08-24🔥🔥

why one ai is better than four 598

08-24🔥

chatgpt logits softmax

08-24🔥

ornith 1 5 9b

08-24🔥

s1 mini

08-23🔥🔥🔥

obliterator qwen38 27b

08-23🔥🔥

qwen3 tts speed cost frontier

08-23🔥🔥

huihui qwen3 8 27b abliterated

08-23🔥🔥

ornith 1 5 35b a3b gguf

08-23🔥🔥

z lab qwen3 8 27b dflash2

08-23🔥🔥

deepseek v4 pro arc agi

08-23🔥🔥

nvidia harness arc agi 3 breakthrough

08-22🔥🔥

nvidia avo arc agi 3

08-22🔥🔥

anthropic multi agent experiment

08-22🔥🔥

anthropic claude mythos 5 enterprise

08-22🔥🔥

google biomarker discovery framework

08-22🔥🔥

artificial analysis mlcr aa benchmark

08-22🔥🔥

nvidia cosmos3 droid lerobot dataset

08-21🔥🔥🔥

welcome to the ai crisis in math

08-21🔥🔥

lfm25 dspark

08-21🔥🔥

multilingual knowledge transfer lexical interventions

08-21🔥🔥

hauhaucs qwen3 8 27b uncensored mtp gguf

08-21🔥🔥

empero qwen38 27b ridge gguf

08-21🔥🔥

ornith 1 5 35b a3b

08-20🔥🔥

glm 5 3 artificial analysis benchmarks

08-20🔥🔥

nvidia cosmos 3 edge robot control

08-20🔥🔥

liquidai lfm25 qad

08-20🔥🔥

glm 5 3

08-20🔥🔥

chatgpt moral judgment

08-20🔥🔥

ntt tsuzumi 2

08-19🔥🔥🔥

qwen 38 27b overthinking

08-19🔥🔥🔥

qwen3 8 2 4t a95b fp8

08-19🔥🔥

qwen3 8 27b scores 52 on artificial analysis

08-19🔥🔥

how much memory does your agent actually need

08-19🔥🔥

microsoft copilot secret input vulnerability

08-19🔥🔥

olmo 3 drug morphology

08-19🔥🔥

deepseek v4 pro 0813 vs gpt 5 6 sol on deepswe

08-18🔥🔥🔥

unsloth qwen3 8 27b nvfp4

08-18🔥🔥

same cluster 33 points more utilization

08-18🔥🔥

three layers of ai agent security

08-18🔥🔥

teaching everyone to fish for tokens

08-18🔥🔥

import ai 469 dig bench faraday

08-18🔥🔥

huggingface state of open models qwen china

08-18🔥🔥

nvidia nemotron rl agentic terminal pivot v1

08-17🔥🔥🔥

nvidia nemotron 3 5 lightning 30b a3b nvfp4

08-17🔥🔥

littlelearner llm pedagogically controlled knowledge exposure

08-17🔥🔥

models are getting dumber on purpose

08-17🔥🔥

inclusionai ling 3 0 tiny model

08-17🔥🔥

prime intellect fable 5

08-17🔥

ulamai unsolved math

08-16🔥🔥

building an ai text detector from scratch

08-16🔥🔥

bytedance seedance 2 5

08-16🔥🔥

qwen 3 8 27b fp8

08-16🔥🔥

muse glimmer 30b gguf

08-16🔥🔥

glm 5 3 cybersecurity ai

08-16🔥

llamaindex extractbench

08-15🔥🔥🔥

qwen qwen3 8 27b

08-15🔥🔥🔥

alibaba opens qwen3 8 max weights

08-15🔥🔥

glm 53 release

08-15🔥🔥

anthropic multi agent conflict

08-15🔥

matraix persona 1m dataset

08-14🔥🔥🔥

introducing gemini 3 7 flash

08-14🔥🔥🔥

deepseek v4 pro release

08-14🔥🔥🔥

sarvam ai opens samvaad to developers

08-14🔥🔥🔥

icml 2200 papers reproduction

08-14🔥🔥

discovered materials benchmark

08-14🔥🔥

perplexity search as code benchmarks

08-14🔥🔥

google gemini 3 7 flash

08-13🔥🔥🔥

spacexai grok 46

08-13🔥🔥🔥

qwen3 8 2 4t a95b

08-13🔥🔥

lm arena claude opus 5 writing style

08-13🔥🔥

upstages solar pro 4 jumps 28 points to beat human agents

08-13🔥🔥

lfm2 5 vl 3b

08-13🔥🔥

recall is the bottleneck for parametric factuality

08-12🔥🔥

stealing reasoning traces from proprietary llm apis

08-12🔥🔥

altk evolve fewer tokens

08-12🔥🔥

mistral regional inference open models

08-12🔥🔥

amie real time clinical video consultations

08-12🔥🔥

nvidia nemotron 3 5 lightning

08-12🔥🔥

openai gpt 56 cyber

08-11🔥🔥🔥

muse glimmer

08-11🔥🔥🔥

anthropic claude riemann zeta

08-11🔥🔥

claude riemann zeta lower bound

08-11🔥🔥

nvidia magpie tts

08-11🔥🔥

deepseek v4 flash debugging benchmark

08-11🔥🔥

making knowledge distillation cheap enough

08-10🔥🔥

lessons from the hacks

08-10🔥🔥

claude opus 5 one week

08-10🔥🔥

ucla finds ai reward hack monitors collapse to 28 on real cheating

08-09🔥🔥🔥

openai slows astra development

08-09🔥🔥

ling 3 0 flash

08-09🔥🔥

liquidai lfm2 5 2 6b

08-09🔥🔥

shieldstral 1 0 3b

08-09🔥🔥

why 99 percent accurate agents fail long horizon tasks

08-09🔥🔥

tutormoments

08-09🔥🔥

anthropic cuts fable 5 biology blocks by 85

08-08🔥🔥🔥

deepseek v4 flash 0731

08-08🔥🔥🔥

openais astra becomes first ai flagged as critically dangerous before release

08-08🔥🔥

tutormoments

08-08🔥🔥

bytedance trains massive ai model

08-08🔥🔥

arbitrage efficient reasoning

08-08🔥🔥

scaling categorical flow maps

08-07🔥🔥

gemini 3 6 flash arc agi

08-07🔥🔥

large genome models virus design

08-07🔥🔥

nvidia alpamayo open model

08-07🔥🔥

openai gpt 5 6 sol

08-07🔥🔥

inclusionai ling 3 0 flash

08-07🔥🔥

nvidia nemotronlabs voicechat 11b

08-06🔥🔥

meta muse code and spark 1 2

08-06🔥🔥

mistral shieldstral 3b safety classifier

08-06🔥🔥

google gemma translator raspberry pi 5

08-06🔥🔥

prime agent beats human experts on arc agi 3

08-06🔥🔥

databricks unity ai gateway ga

08-06🔥🔥

artificial analysis endpoint accuracy index

08-04🔥🔥🔥

alibaba releases qwen max open weight ai

08-04🔥🔥

openai s astra cracks 10 math problems

08-04🔥🔥

epoch mirrorcode claude fable 5

08-04🔥🔥

openai rebuilds chatgpt voice with gpt live to talk and listen at once

08-04🔥🔥

qwen3 8 max

08-04🔥🔥

microsoft orchard framework

08-04🔥🔥

self sustaining ai viruses and pacing ai progress

08-04🔥🔥

claude opus 5 debugging benchmark

08-03🔥🔥🔥

openai astra math

08-03🔥🔥

latest open artifacts 23

08-03🔥🔥

xyz ailab xyz aquila mini

08-03🔥🔥

qyrou reasoning corpus 4k 5m v1

08-03🔥🔥

thinkingmachines inkling small

08-02🔥🔥🔥

kimi k3

08-02🔥🔥

is ai reasoning right for the wrong reasons

08-02🔥🔥

claude opus 5 vending bench 2

08-02🔥🔥

luffythefox qwen36 35b genesis

08-02🔥🔥

microsoft fara1.5 27b

08-02🔥🔥

inflect micro v2

08-02🔥🔥

skywork ai mureka v9

08-01🔥🔥🔥

epoch expands frontiermath to 50 unsolved problems ai has already cracked three

08-01🔥🔥

oxide and friends open weight revolution simon willison

08-01🔥🔥

open model race four ways

08-01🔥🔥

alibaba qwen audio 3 0 asr flash

08-01🔥🔥

deepseek v4 flash agent benchmarks

08-01🔥🔥

minimax h3 video editing

2026 · Jul

07-31🔥🔥🔥

gemini robotics er 2

07-31🔥🔥

thinking machines inkling small

07-31🔥🔥

echoverse deep evolving environments for computer use agents

07-31🔥🔥

goodfires silico catches qwen3 35b endorsing drunk driving

07-31🔥🔥

sarvam ai saaras v4 asr model

07-31🔥🔥

ai adoption j curve

07-31🔥🔥

sarvam ai ships bulbul v4 to bring emotional voice to 11 indian languages

07-30🔥🔥

handbook md long policy documents agents

07-30🔥🔥

claude opus 5 vending bench

07-30🔥🔥

sakana ai dream cubed

07-30🔥🔥

tencent open sources angelspec

07-30🔥🔥

goodfire silico fixes ai training collapse

07-29🔥🔥🔥

discovering cryptographic weaknesses with claude

07-29🔥🔥🔥

benchmarking opus 5 on slop code bench

07-29🔥🔥🔥

kimi k3 weights

07-29🔥🔥

discovering cryptographic weaknesses with claude

07-29🔥🔥

googles data shows workers arent automating themselves away

07-29🔥🔥

claude opus 5 deepswe

07-28🔥🔥🔥

moonshot ai kimi k3 2 8t model

07-28🔥🔥

claude opus 5 release

07-28🔥🔥

goodfires silico finds human cognitive biases geometrically encoded inside llms

07-28🔥🔥

nvidia ising calibration 1 5

07-28🔥🔥

nvidia cosmos h dreams surgical robotics

07-28🔥🔥

gh esd grounded hypothesis driven error slice discovery

07-28🔥🔥

qwen38 max distillation 50k

07-27🔥🔥

claude opus 5 release

07-27🔥

how to evaluate a new ai model without starting from scratch

07-26🔥🔥🔥

anthropic claude opus 5 release

07-26🔥🔥🔥

anthropic claude opus 5

07-26🔥🔥

arc agi leaderboard

07-26🔥🔥

quoting boris cherny

07-26🔥🔥

who gets to understand ai

07-26🔥🔥

qwen3.6 27b fable fusion 711

07-26🔥🔥

motif 3 beta release

07-26🔥🔥

running a 289m parameter llm on an 8 microcontroller

07-26🔥🔥

kwaipilot kat coder v2 5 dev

07-26🔥🔥

minicpm robotmanip

07-26🔥🔥

poolside laguna s 2.1

07-26🔥🔥

upstage solar open2 250b

07-26🔥

nanbeige nanbeige42 3b

07-25🔥🔥🔥

claude opus 5

07-25🔥🔥

flux 3 mimic

07-25🔥🔥

lead breaking the no recovery bottleneck

07-25🔥🔥

glint research fable 5 traces

07-25🔥🔥

claude fable 5 traces

07-24🔥🔥

amd and cerebras pair helios with wafer scale engine for 5x faster inference

07-24🔥🔥

context rot is a harness problem and rlms show why

07-24🔥🔥

browser use ships game mode

07-23🔥🔥🔥

introducing gemini 3 6 flash 3 5 flash lite and 3 5 flash cyber

07-23🔥🔥🔥

openai sandbox escape

07-23🔥🔥

kimi k3 fable

07-23🔥🔥

symptomai conversational ai agent symptom assessment

07-22🔥🔥🔥

gemini 3 6 flash 3 5 flash lite and 3 5 flash cyber

07-22🔥🔥🔥

gpt 5 6 sol benchmark hack

07-21🔥🔥

how we measured ai writing on arxiv

07-21🔥🔥

length value model

07-20🔥🔥🔥

gpt 5 6 claude fable 5 muse spark 1 1

07-20🔥🔥

angel slim hy3 gguf

07-20🔥🔥

minicpm5 1b claude opus fable5 thinking

07-20🔥🔥

wan dancer 14b

07-19🔥🔥🔥

kimi k3 28t a50b

07-19🔥🔥

fable 5 vs gpt 5 6 sol np hard

07-19🔥🔥

controlling reasoning effort in llms

07-19🔥🔥

ovisocr2

07-18🔥🔥

schema harness arc agi 3

07-18🔥🔥

newer models same advantage

07-17🔥🔥🔥

moonshot ai kimi k3

07-17🔥🔥🔥

openai gpt red automated red teaming

07-17🔥🔥

can llms perform deep technical comprehension of computer architecture papers

07-17🔥🔥

nvidia nemotron 3 embed

07-16🔥🔥🔥

mira muratis thinking machines drops inkling

07-16🔥🔥

what building shippy taught us about building agents

07-16🔥🔥

model routing is simple until it isnt

07-16🔥🔥

exploring hierarchical interest representation for meta ads deep funnel optimization

07-16🔥🔥

introducing real world voiceeq

07-16🔥🔥

clara bridging retrieval and generation

07-16🔥🔥

uncertainty quantification for llm function calling

07-15🔥🔥

coding agents think ahead

07-15🔥🔥

nemotron labs open models

07-15🔥🔥

empowering indias next generation of innovators with atl saathi

07-15🔥🔥

proactive agent research environment

07-15🔥🔥

moss transcribe diarize

07-15🔥🔥

nvidia nemotron labs audex 30b a3b

07-15🔥🔥

unsloth deepseek v4 flash gguf

07-15🔥🔥

glm 5 2 release

07-13🔥🔥

global workspace llm

07-11🔥🔥

gpt 5 6 grok 4 5 claude muse spark build off

07-11🔥🔥

incentivizing temporal awareness in egocentric video understanding models

07-11🔥

glm 5 2 vat benchmark

07-11🔥

building a real time ai tutor for 5 year olds

07-10🔥🔥🔥

gpt 5 6 family

07-10🔥🔥

grok 4.5 gpt 5.5 claude build off

07-10🔥🔥

glm 5 2 vat benchmark

07-10🔥🔥

sensorfm towards a general intelligence and interface for wearable health data

07-10🔥

data for agents

07-09🔥🔥

expanding managed agents in gemini api

07-09🔥🔥

nvidia nemotron langchain deep agents

07-09🔥🔥

flint visualization language

07-08🔥🔥

intelligence is free now what

07-08🔥🔥

lerobot v0 6 0

07-08🔥🔥

huggingface kernels update

07-08🔥

prx part 4 data strategy

07-07🔥🔥

does code cleanliness affect coding agents

07-07🔥🔥

fable 5 vending bench

07-06🔥🔥

the log is the agent

07-06🔥🔥

hf trending deepseek v4 pro dspark

07-04🔥

amortizing maximum inner product search

07-03🔥

2026 bair graduate showcase

07-03🔥

google ai updates june 2026

07-02🔥🔥

how nvidias inference software stack powers the lowest token cost

07-02🔥🔥

why specialization is inevitable

07-02🔥🔥

hugging face cerebras gemma4 voice ai

07-02🔥🔥

scarfbench

07-01🔥🔥

claude meets blackwell ultra

07-01🔥🔥

start building with nano banana 2 lite and gemini omni flash

07-01🔥

words are a byproduct of consciousness

2026 · Jun

06-30🔥

tokenmaxxing is dead long live tokenmaxxing

06-30🔥

the cost yagni was never about

06-29🔥🔥🔥

gpt 5 6 sol deepseek dspark

06-29🔥🔥

glm 5 2 beats claude

06-27🔥🔥🔥

quoting openai

06-26🔥🔥

which tokens does a hybrid model predict better

06-26🔥🔥

accelerating transformers fine tuning with nvidia nemo automodel

06-26🔥🔥

understanding the brain with ai driven explanations and experiments

06-25🔥🔥🔥

google gemini 3 5 flash computer use

06-25🔥🔥

qwen agentworldbench

06-24🔥🔥🔥

openai daybreak gpt 5 5 cyber

06-24🔥🔥🔥

daybreak securing the world

06-24🔥🔥🔥

plamo 3 0 prime

06-24🔥🔥

unlimited ocr

06-24🔥🔥

2026 06 24 papers role confusion 2025

06-24🔥🔥

2026 06 24 papers 20260622 role confusion injection

06-23🔥🔥

glm 5 2 open weights release

06-22🔥🔥

softmax free attention gpt2 medium

06-22🔥

qwopus 3 6 27b coder

06-21🔥

accessing books3 dataset for research

06-20🔥🔥🔥

deepseek v4 efficient million token

06-20🔥🔥

huggingface peft beyond lora benchmark

06-20🔥🔥

openai diagnose rare genetic diseases

06-20🔥🔥

claude opus 4 8 quality regression analysis

06-19🔥🔥🔥

openai ai chemist medicinal chemistry

06-19🔥🔥🔥

glm 5 2 open weights model

06-19🔥🔥

from the hugging face hub to robot hardware with strands agents and lerobot

06-19🔥🔥

is it agentic enough

06-18🔥🔥🔥

nvidia blackwell mlperf training 6 0

06-18🔥🔥

amie medical ai disease management

06-18🔥🔥

unlocking uk house building with ai accelerated planning

06-18🔥🔥

agentic resource discovery

06-18🔥🔥

molmomotion

06-15🔥🔥

rio llm merge controversy

06-15🔥🔥

google gemma 4 12b unified encoder free

06-14🔥🔥

cohere north mini code 1 0

06-14🔥🔥

minimax m3

06-14🔥🔥

ai2 olmo eval workbench

06-13🔥🔥

gemma 4 12b it

06-12🔥🔥

llms tactical nukes simulation

06-11🔥🔥

nvidia accelerates diffusiongemma

06-11🔥🔥

diffusiongemma 4x faster text generation

06-10🔥🔥

cohere north mini code

06-10🔥🔥

servicenow bilingual asr benchmark

06-08🔥🔥🔥

2026 06 08 papers 2412.19437

06-08🔥🔥

tokenomics agentic software engineering

06-08🔥🔥

coheres unreleased coding model

06-08🔥

llm human attributes aoe2

06-07🔥🔥

google magenta mrt2 low latency

06-06🔥🔥

thousand token wood 3b agent economy

06-06🔥

2026 06 06 papers 2606 04037

06-05🔥🔥

dpo beyond chatbots

06-04🔥🔥🔥

microsoft mai thinking 1 moe

06-03🔥🔥

holo31 fast local computer use

06-02🔥🔥

jetbrains mellum2 12b moe

06-02🔥🔥

nvidia nemotron 3 ultra

06-01🔥🔥🔥

anthropic claude opus 4 8 release

06-01🔥🔥

qwopus 3 6 27b v2 mtp gguf

2026 · May

05-31🔥🔥

parallax parameterized local linear attention

05-31🔥🔥

probe targeted fine tuning

05-30🔥

a moment of thanks for deepseek

05-29🔥🔥🔥

laguna m1 xs2

05-29🔥🔥

itbench aa frontier models score below 50 percent

05-29🔥🔥

2026 05 29 papers 2605.27375

05-25🔥🔥

nvidia nemotron labs diffusion

05-25🔥🔥

psibotai syndata

05-24🔥🔥🔥

cohere command a plus 218b moe

05-23🔥

2026 05 23 papers 2605.20189 solar lifelong learning

05-22🔥🔥🔥

google gemini 3 5 flash vs all

05-22🔥🔥

gemini system prompt leak

05-21🔥🔥

huggingface ettin reranker family release

05-21🔥🔥

internlm intern s2 preview 35b

05-20🔥🔥🔥

google gemini 3 5 flash agentic performance

05-20🔥🔥🔥

gemini 3 5 flash

05-19🔥🔥🔥

transformer scalability crisis

05-19🔥🔥🔥

nvidia nemotron personas korea

05-19🔥🔥🔥

qwen 3 7 dropped on qwen chat

05-17🔥🔥

llm architectures kv sharing mhc

05-17🔥🔥

arxiv llm error ban

05-16🔥🔥🔥

teichai deepseek v4 pro agent dataset

05-15🔥🔥🔥

inclusionai ring 2 6 1t

05-15🔥🔥

hermes agent reasoning traces

05-14🔥🔥🔥

mimo v25 pro opensource

05-14🔥🔥

modotte codex 2m thinking

05-13🔥🔥🔥

jina embeddings v5 omni

05-11🔥🔥

llms corrupt documents delegation

05-11🔥🔥

hy mt 1 5 1 8b 1 25bit quantization

05-11🔥🔥

hidream o1 image uit

05-10🔥🔥

teaching claude why

05-09🔥🔥

allen institute emo moe modularity

05-09🔥🔥

cybersecqwen 4b

05-09🔥🔥

ai2 emo moe

05-08🔥🔥🔥

openai voice intelligence api

05-08🔥🔥🔥

gpt 5 5 and cyber trusted access

05-08🔥🔥

natural language autoencoders

05-06🔥🔥🔥

gpt 5 5 instant system card

05-06🔥

microsoft nsdi 2026 advances

05-05🔥🔥

inclusionai ling 2 6 flash release

05-05🔥🔥

qwen3 6 27b dflash speculative decoding

05-05🔥🔥

autobe benchmark backend generation

05-04🔥🔥🔥

kimi k2 6 beats gpt 5 5 coding

05-04🔥🔥

harvard o1 er diagnosis

05-04🔥🔥

evolving deep learning optimizers

05-03🔥🔥

refusal in language models is mediated by a single direction

05-03🔥🔥

deepseek v4 flash

05-03🔥🔥

nvidia nemotron 3 nano omni 30b a3b reasoning bf16

05-03🔥🔥

unsloth qwen3 6 27b gguf

05-02🔥🔥🔥

fineweb edu

05-02🔥🔥

deepseek v4 series release

05-02🔥🔥

ai outperforms er doctors diagnostic cases

05-02🔥🔥

grok 4 3 benchmark performance

05-02🔥🔥

llm refusal single direction

05-02🔥🔥

xiaomimimo mimo v2 5

05-02🔥🔥

xiaomimimo mimo v2.5 pro

05-01🔥🔥

qwen3 6 27b uncensored hauhaucs aggressive

05-01🔥🔥

gpt 55 cyber capabilities

05-01🔥🔥

red teaming a network of agents

2026 · Apr

tools (287)

2026 · Sep

2026 · Aug

08-31🔥🔥

jetbrains junie local

08-31🔥

agent skill token cost

08-31🔥

jsonl log data governance

08-31🔥

apache 2 0 ai model compliance

08-30🔥🔥

vllm v0280 release

08-29🔥🔥

perplexity search api

08-29🔥🔥

stripes link wallet lets ai agents buy things

08-29🔥🔥

exa dynamic highlights

08-28🔥🔥

serve markdown to ai agents with accept headers

08-28🔥🔥

experientiallabs experiential

08-28🔥🔥

run claude managed agents with chat sdk

08-28🔥🔥

pi coding agent

08-28🔥🔥

vllm v0 28 0 sparse attention speculative decoding

08-28🔥🔥

ai engineer notebooks

08-27🔥🔥🔥

huihui ai qwen3 8 27b abliterated gguf

08-27🔥🔥

rag is simpler than you think

08-26🔥🔥🔥

ox alpha

08-26🔥🔥

llms could control their host machines by exploiting inference engines

08-26🔥🔥

headlong microharness

08-26🔥🔥

how to evaluate llms before production

08-26🔥🔥

nvidia dynamo shadow engine recovery

08-26🔥🔥

agentic observability with amazon opensearch service mcp apps

08-26🔥🔥

coding agent harness source code analysis

08-25🔥🔥

kimi k3 glm 5 exploit

08-25🔥🔥

fabien sanglard agent md

08-25🔥🔥

ocr it chrome extension

08-24🔥🔥

why your local llm feels dumber than it is

08-24🔥🔥

qwen 3 8 27b reverse engineering

08-24🔥🔥

llm 0 33

08-23🔥🔥

model context protocol roadmap 2026

08-23🔥🔥

claude code shorts orchestration

08-23🔥🔥

codex aws bedrock bug

08-23🔥🔥

latentspace ai simulation trend

08-23🔥🔥

the evolution of the agent harness

08-22🔥🔥

ox alpha

08-22🔥🔥

reduce rag costs on amazon bedrock

08-22🔥🔥

claude code subagent patterns

08-22🔥🔥

resource control local ai gateway

08-22🔥🔥

ai continuous generation drift prevention

08-21🔥🔥

ramp launches ai model router

08-21🔥🔥

introducing cross region inference for openai gpt 5 6 models on amazon bedrock

08-18🔥🔥

moneyforward ai token management

08-18🔥🔥

cloudflare ai search rag

08-17🔥

design ai systems with provisional agents

08-17🔥

microsoft agent framework for go

08-17🔥

ai article pipeline human gate

08-17🔥

knowledge graph agent

08-16🔥🔥🔥

unsloth qwen3.8 27b gguf

08-16🔥🔥

thoughtdag

08-16🔥🔥

vllm dspark deepseek v4

08-16🔥🔥

nvidias nemo switchyard cuts agent ai costs by 74

08-15🔥🔥

warp brings xai grok 4 6

08-15🔥🔥

nous research hermes agent loop command

08-15🔥🔥

maximizing claude code sessions

08-15🔥🔥

gemini 3 7 flash

08-15🔥🔥

lightricks ltx 2.5 video generation model

08-14🔥🔥🔥

cerebras openai gpt 5 6 sol ultrafast

08-14🔥🔥

strands agents lerobot hugging face

08-14🔥🔥

artificial analysis optima

08-14🔥🔥

ai agent nenrin

08-13🔥🔥

llama cpp

08-13🔥🔥

shieldfont ai scrapers

08-13🔥🔥

nous research hermes agent profiles

08-13🔥🔥

oneadvanced deployed over 50 ai agents on uk sovereign aws

08-12🔥🔥🔥

meta models muse glimmer 30b

08-12🔥🔥

programming language token cost evaluation

08-12🔥🔥

amazon bedrock openai daybreak

08-12🔥🔥

onestruction ishigaki ids

08-11🔥🔥

openchamber

08-11🔥🔥

sarvam ai speech to text api dezerv

08-10🔥🔥🔥

agent plugins 100

08-10🔥🔥

claudecode cross session messaging

08-10🔥🔥

orca parallel agents

08-10🔥🔥

ai review findings verification

08-09🔥🔥

transform cost mifos

08-09🔥🔥

anthropic managed agents budget caps

08-09🔥🔥

huggingfaceh4 ultrachat 200k

08-08🔥🔥

prime intellect multi agent rl training

08-08🔥🔥

cloudflare radar researcher

08-07🔥🔥🔥

openai microsoft and cursor unite behind agent plugins

08-07🔥🔥

perplexity swaps gpt 5 6 terra and luna into its ai agent platform

08-06🔥🔥

design md measured

08-04🔥🔥🔥

the inference engineering masterclass

08-04🔥🔥

jfrog uncovers fake sqlite cves llm slop

08-04🔥🔥

llms reward expertise

08-04🔥🔥

airllm 70b inference single 4gb gpu

08-03🔥🔥

wafer kimi k3 mi355x

08-03🔥🔥

unsloth kimi k3 gguf

08-03🔥🔥

welcome to agents week

08-03🔥🔥

rpr ai governance runtime

08-03🔥🔥

deepseek v4 flash api

08-02🔥🔥

waste kimi k3 engine

08-02🔥🔥

the session you cannot take with you

08-02🔥🔥

datasette apps 0 2a0

08-02🔥🔥

unsloth deepseek v4 flash 0731 gguf

08-02🔥🔥

wordpress mcp server

08-01🔥🔥

deepseek v4 flash public beta

08-01🔥🔥

poolside laguna s 2 1 rate limits

08-01🔥🔥

mac laguna ollama llama cpp vllm

2026 · Jul

07-31🔥🔥🔥

introducing explicit prompt caching for openai gpt 5 6 models on amazon bedrock

07-31🔥🔥

claude shared chat leak

07-31🔥🔥

claude api structured outputs compaction adaptive thinking

07-31🔥🔥

deploying kimi k3 on aws

07-30🔥🔥🔥

document borne ai worms copilot word

07-30🔥🔥🔥

how agentcore gateway supports the mcp 2026 07 28 spec

07-30🔥🔥

self hosting kimi k3

07-30🔥🔥

moonshotai kimi k3

07-30🔥🔥

how to self host a validated ai coding assistant with nvidia nemo guardrails

07-30🔥🔥

sub2api

07-29🔥🔥

gemini api managed agents 3 6 flash hooks

07-29🔥🔥

kimi k3 telnyx inference

07-28🔥🔥🔥

prism ml ternary bonsai 27b gguf

07-28🔥🔥

microsoft security ai tools

07-28🔥🔥

beyond rag task aware knowledge compression for enterprise ai on aws

07-28🔥🔥

moonshot ai open sources flashkda

07-28🔥🔥

turboquant kv cache quantization

07-27🔥🔥🔥

prism ml bonsai 27b gguf

07-27🔥🔥

relay market investigation

07-27🔥🔥

vllm v0 26 0 inkling 1t

07-27🔥🔥

mcp 2026 07 28 stateless

07-27🔥🔥

poolside laguna s 21 nvfp4

07-27🔥🔥

unsloth laguna s 2 1 gguf

07-26🔥🔥🔥

empero ai qwythos 9b claude mythos 5 1m gguf

07-26🔥🔥

claude cookbook

07-26🔥🔥

hetzner inference experiment

07-26🔥🔥

two dimensional framework for ai agent design patterns

07-26🔥🔥

opus5 breaking changes prompts

07-26🔥🔥

huggingfacecode stack v3 train

07-26🔥🔥

poolside laguna s 2.1 gguf

07-25🔥🔥

get started with openai gpt 5 6 sol terra and luna on amazon bedrock

07-24🔥🔥

echo model router

07-24🔥🔥

cactus hybrid

07-24🔥

i regret migrating to codeberg

07-23🔥🔥

gigatoken fast tokenization

07-23🔥🔥

laguna s 2 1

07-23🔥

drawing the mona lisa with gpt 5 6 claude gemini and grok

07-22🔥🔥

agent swarms and the new model economics

07-21🔥🔥

nativ mac local model runner

07-21🔥🔥

amazon quick nvidia nemo agent toolkit

07-20🔥🔥

openai reduces codex model context size

07-18🔥🔥

claude fable 5 gpt 5 6 sol music video arena

07-18🔥🔥

lm studio bionic

07-18🔥🔥

how smartsheet built a remote mcp server on aws

07-17🔥🔥

detecting llm generated texts with classical machine learning

07-17🔥

the llm critics are right i use llms anyway

07-16🔥🔥

i tricked claude into leaking your deepest darkest secrets

07-16🔥🔥

towards a harness that can do anything

07-16🔥

running gemma 4 26b on xeon

07-15🔥🔥

how to stop claude from saying load bearing

07-15🔥🔥

rl agent trains ai

07-15🔥

the agentic loop

07-14🔥🔥

add flag for ai generated articles

07-14🔥🔥

i love llms i hate hype

07-14🔥🔥

migrating a production ai agent to gpt 5 6

07-13🔥🔥🔥

claude code vs opencode token overhead

07-13🔥🔥

old and new apps via modern coding agents

07-12🔥🔥

deploying quantized models on amazon sagemaker ai with unsloth

07-12🔥🔥

better tools made copilot code review worse

07-11🔥🔥

introducing muse spark 1 1

07-10🔥🔥🔥

gpt 5 6 sol luna terra vercel

07-10🔥

i think i have llm burnout

07-10🔥

colibri glm 5 2 runner

07-09🔥🔥

build a serverless image editing agent with amazon bedrock agentcore harness

07-09🔥🔥

native speed vllm transformers backend

07-09🔥🔥

tencent hy3

07-09🔥🔥

automating cross repo documentation with github agentic workflows

07-07🔥🔥

qwen agentworld 35b a3b

07-06🔥🔥🔥

claude fable 5 autonomous agent

07-06🔥🔥

internscience agents a1

07-05🔥🔥

llm coding ux experimentation

07-04🔥🔥

jamesobs guide to running sota llms locally

07-04🔥🔥

pxpipe cost reduction

07-04🔥

memorizing session transcripts

07-03🔥🔥

kimi k2 7 code github copilot

07-03🔥🔥

senior swe bench

07-03🔥🔥

no llm code in dependencies

07-02🔥🔥

zcode claude code from the makers of glm

07-02🔥🔥

claude fable 5 promotional access

07-02🔥🔥

zcode harness for glm 5 2

07-01🔥🔥

claude sonnet 5

07-01🔥🔥

claude science

2026 · Jun

2026 · May

2026 · Apr

business (102)

2026 · Sep

2026 · Aug

08-31🔥🔥🔥

openai cuts cursor s access to gpt models after spacex takeover

08-31🔥🔥

metr redwood huggingface hack postmortem

08-30🔥🔥🔥

openai terminates cursor model provision

08-30🔥🔥🔥

sony music warner sue anthropic

08-30🔥🔥🔥

open weight ai companies acquisition boom

08-28🔥🔥🔥

open executive

08-28🔥🔥🔥

anthropic supply chain risk lawsuit judge ruling

08-28🔥🔥

openai agents unauthorized hugging face incursion

08-28🔥🔥

anthropic slashes claude team pricing 85 for academic researchers

08-26🔥🔥

openrouter china model share 58 percent

08-26🔥🔥

ox alpha rumors

08-25🔥🔥

openai gpt 5 6 sol price reduction

08-25🔥🔥

openai is building ai agents for everything

08-24🔥🔥

openai gpt 5 6 sol api pricing

08-24🔥🔥

openai safety process adjustment

08-24🔥🔥

anthropic ai model revenue ramp index

08-23🔥🔥

inherent faraday ai agent research replication

08-22🔥🔥

moonshot ai kimi japan expansion

08-21🔥🔥🔥

stripe openrouter acquisition

08-21🔥🔥

unitree ceo robotics chatgpt moment

08-20🔥🔥🔥

anthropic 65b samsung 15 wrc2026

08-20🔥🔥🔥

openai halts frontier ai rl security

08-20🔥🔥

cognition ceo denies spacex acquisition report

08-19🔥🔥🔥

stripe buys openrouter for 7b

08-18🔥🔥

caddi manufacturing os

08-18🔥🔥

amazon rare books ai training

08-17🔥🔥

debian begins voting on llm contributions

08-17🔥

why people arent buying mark zuckerbergs ai future

08-16🔥🔥🔥

metr raises 71m to independently stress test ai

08-16🔥🔥

openai and anthropic price war

08-15🔥🔥🔥

spacex closes 60b cursor deal

08-15🔥🔥

deepseek peak off peak pricing update

08-15🔥🔥

openai hires new cro dali rajic

08-14🔥🔥🔥

databricks raises 5b at 190b valuation

08-14🔥🔥

deepseek api pricing update

08-13🔥🔥🔥

lovable confirms 13 3b valuation raises 400m

08-13🔥🔥

cerebras cloud revenue

08-12🔥🔥🔥

general catalyst leads 1 1b round into 2 month old river ai

08-11🔥🔥

meta open models reboot

08-10🔥🔥🔥

ai safety evaluation sandbox escape incidents

08-09🔥🔥

anthropic fable 5 biology guardrails 緩和

08-08🔥🔥

deepseek api price hike

08-07🔥🔥🔥

amd acquires taalas

08-07🔥🔥

anthropic confirms plans to build an in house silicon team

08-04🔥🔥

openai super pac funding ai news site

08-02🔥🔥🔥

openai academic researchers program

08-01🔥🔥🔥

openai gpt 5 6 luna price cut

08-01🔥🔥

google fixed more chrome bugs with ai

2026 · Jul

2026 · Jun

2026 · May

2026 · Apr