right when i was thinking of canceling my oai sub, they drop o3-mini (which is apparently super good with ml stuff)
when o3-mini in perplexity @AravSrinivas? pplx pro has replaced chatgpt for me since the past few weeks
@huggingface@deepseek_ai@asquirous
Was using R1 on HuggingChat
Model randomly started generating the reasoning trace in Mandarin.
Even the output was completely in Mandarin.
@knowrohit07@asquirous Would it not be possible to get a statistical measure of perceived DoF over large number of sample runs with fixed initial settings?
The exact working of networks is not known, but I'd assume they show a trend over a large enough sample space.
@knowrohit07@asquirous Given a fixed set of initial conditions, I imagine it should be possible to determine the perceived DoF that a Diffusion model provides.
The only hitch is the network is a black box in terms of interpretability, so there would be some randomness even in a controlled setting.
Quite exxcited to announce the AI Starter Pack! π₯
If you've been thinking about getting into AI - this is cue - get 1000 of dollars to try out the bleeding edge of AI β‘
End the year with a bang! - APPLY today! π₯
Using the SLM to get a speculative (approximate) distribution, which can then be refined by the LLM also limits sampling overhead, something faced by regular decoding due to the large output token distribution of LLMs.
This illustration of token sampling from @JayAlammar's llm-book really puts the idea of speculative decoding into perspective.
Using an SLM and LLM with similar output probability distributions allows for refinement of the final output probability distribution.
@HiHarshSinghal Would be cool to have a cafe called "The Latent Space". It could have a couple of digital screens on the walls, and people walk up to them to generate art through prompts.
@reach_vb Definitely.
A platform wide HF chatbot which indexes both model / dataset cards + majorly used HF documentation would make using / traversing HF easier.