Top Tweets for #InFerence
L’inférence IA à moitié prix avec OpenRouter Batch 💸
Pourquoi continuer à payer le prix fort pour des tâches d’IA qui n’ont pas besoin d’une réponse instantanée ?
#KingLand #IA #OpenRouter #Inference #SaaS #Cloud #DataEngineering #CTO #Productivite #IntelligenceArtificielle #Innovation #Technologie
▫️ Fiche Impact : https://t.co/wRG93JKYEf
Parce qu’avec les traitements asynchrones, la patience vaut mieux que la ruine.
Alors que la pression budgétaire s’intensifie sur les infrastructures d’intelligence artificielle, l’exécution en temps réel devient un luxe évitable. La classification de données, l’enrichissement de catalogues ou l’évaluation de modèles n’ont aucun besoin d’une interactivité immédiate. En unifiant les API Batch de plus de 70 modèles sous une interface commune, @OpenRouter propose une alternative majeure pour rationaliser les dépenses d’inférence d’arrière-plan.
📦 Une seule API pour gouverner vos lots : Centralisez vos requêtes pour des dizaines de modèles sans multiplier les intégrations spécifiques.
💰 Des tarifs réduits de moitié : Profitez d’une baisse moyenne de 50 % sur le coût par token pour les traitements tolérant un délai.
⚡ Une intégration simplifiée par JSON : Envoyez vos requêtes directement dans un tableau structuré, sans avoir à gérer le téléversement de fichiers JSONL lourds.
🛠️ Le support du Bring Your Own Key : Conservez vos accords tarifaires avec vos fournisseurs tout en déléguant la file d’attente à OpenRouter.
"La véritable maturité de l’ingénierie IA ne se mesure pas à la rapidité brute d’un prompt, mais à l’intelligence économique de son routage. Savoir ralentir ce qui peut l’être pour financer ce qui doit briller en direct est le nouveau super-pouvoir des architectes cloud."
— C. Pestel
En observant les architectures de nos partenaires, je remarque souvent une obsession stérile pour la milliseconde, même sur des tâches de fond qui tournent la nuit. Cette nouveauté met en lumière un basculement de paradigme : la latence n’est plus seulement une contrainte technique, elle devient une monnaie d’échange. Pouvoir négocier le temps contre du budget ouvre enfin la voie à des agents IA viables à l’échelle, sans risquer l’asphyxie financière à la première hausse d’activité.
📇 Fiche Tool : https://t.co/huqXzEplsa
Quelles tâches de vos pipelines actuels pourriez-vous basculer en différé dès ce soir pour diviser votre facture par deux ?

University of Manchester is now running NVIDIA's Earth-2 model to forecast UK air pollution — hourly particulate-matter maps refreshed every 15 minutes, a tenfold jump in temporal granularity over the chemistry simulators it replaces. https://t.co/T543zohWOC #Inference #NVIDIA

@rohanpaul_ai If smaller DeepSeek models really sit on gaming GPUs, that’s a cost curve story. #DeepSeek #Inference #GPUs
Imagine a conversation with an AI so fluid and fast that the technology becomes completely invisible. The momentum Nicole Junkermann identified at Grow continues to set new standards.
#Grow #RealTimeAI #Hardware #Inference #NicoleJunkermann

@BlackRock @XDCNetwork @xdcaitech @atulkhekade @CoinbaseDev @circle 4/7 🧵 Then #AICompute. @BlackRock says compute may become a programmable resource agents discover, provision + pay for. DiCompute is building decentralized inference with XDC settlement designed into its stack. its on-chain $XDC settlement is not live yet. #XDC #DePIN #Inference
#statstab #623 Six Lemmas Concerning Heterogeneity and Nonlinearity in Causal Inference
Thoughts: A very clear treaty of Causal inference problems.
#causalinference #ATE #MachineLearning #inference #bias #heterogeneity #lemma #nonlinear #observational
https://t.co/twJT7yISOy
How to Follow a Reading Passage
https://t.co/9BIG0q64tJ
#ged #hiset #hse #adulteducation #reading #readingcomprehension #mainidea #detail #inference #vocabulary #context #literacy

Are you stuck with #Anthropic for #AI inference?
I successfully moved our #inference and other workload to AWS, and also received AWS credits upto 25% of my current anthropic billing. Not sure how to do it? DM me and I'll help you out
#claude #aws #opus #jev #cloud
Anthropic and OpenAI launch Opus 5.5 and GPT-6 Sol, slashing frontier AI inference costs for enterprise agents.
#anthropic #openai #llm #inference
TypeSafe AI debuts Jev, a native decision model that replaces text generation with probabilities to cut LLM routing costs and latency.
#llm #inference #decisionmodels #systemarchitecture
SiliconFlow has raised roughly $433M in 2026 as demand grows for cheaper, faster model inference. Training gets the headlines, but inference economics determine how expensive every useful AI interaction becomes after deployment.
#SiliconFlow #AIInfrastructure #Inference
@AethirEco @AethirCloud @nvidia @AxeCompute 4/6 🧵 Scale matters. #Aethir reports 430,000+ #GPU containers across 200+ locations in 94 countries while charging zero data-egress fees. For #Inference-heavy workloads, that attacks 2 pain points at once: #ComputeAvailability + #Data-movement cost. ⚙️ #GPUaaS #Compute #ATH
AI’s next speed boost may be software, not silicon. Smarter scheduling and caching could unlock today’s chips.
#AI #Inference #Software https://t.co/f9wScVzD6J
$X2M would be better off doing a tie-up with $DXN $DXN.ax could be a deadly combination for distributed Data centres, landing points, achoring stations.
Distributed, modular small scale DCs are going to see quite a bit of demand to service #inference
Disc Holding $X2M.ax
AWS SageMaker AI re-architects its inference stack with cache-aware routing and tiered KV caching to boost generative AI throughput.
#machinelearning #inference #awssagemaker #mlops
Euclyd raises $231M to build on-premises AI inference silicon, challenging cloud GPU dominance for enterprise workloads.
#aihardware #inference #semiconductors #venturecapital
Imagine a conversation with an AI so fluid and fast that the technology becomes completely invisible. The momentum Nicole Junkermann identified at Grow continues to set new standards.
#Grow #RealTimeAI #Hardware #Inference #NicoleJunkermann

@nvidia @huggingface @AethirCloud 3/4🧵 Why this matters: if model discovery stays open, compute becomes the next battleground. $ATH targets decentralized #GPU supply for inference. No direct partnership was announced; this is a market thesis, not deal participation.⚡🖥️🛰️🎯 #DePIN #Inference #GPUCloud #AICompute
Last Seen Hashtags on Sotwe
minichat
Teenagegirls filter:videos
Seen from United States
minichat((((((()))))))* filter:native_video
incesto padre y hija
Seen from Argentina
アズールレーン
Seen from Japan
زبة
Seen from Algeria
thundr((((((()))))))* filter:native_video
Seen from Japan
29รับ100
Seen from Thailand
zoophillia
Seen from Netherlands
Trends for you
Most Popular Users

Elon Musk 
@elonmusk
241.7M followers

Barack Obama 
@barackobama
119M followers

Cristiano Ronaldo 
@cristiano
114.4M followers

Donald J. Trump 
@realdonaldtrump
111.9M followers

Narendra Modi 
@narendramodi
107.2M followers

Rihanna 
@rihanna
98.7M followers

NASA 
@nasa
92.4M followers

Justin Bieber 
@justinbieber
91.8M followers

KATY PERRY 
@katyperry
90M followers

Taylor Swift 
@taylorswift13
83.9M followers

Lady Gaga 
@ladygaga
75.4M followers

Virat Kohli 
@imvkohli
73.3M followers

Kim Kardashian 
@kimkardashian
70.9M followers

YouTube 
@youtube
68.8M followers

Neymar Jr 
@neymarjr
66.4M followers

Bill Gates 
@billgates
65.2M followers

Selena Gomez 
@selenagomez
63M followers

The Ellen Show
@theellenshow
62.3M followers

CNN 
@cnn
61.8M followers

X 
@x
60.7M followers















