home.social

#phindinstant β€” Public Fediverse posts

Live and recent posts from across the Fediverse tagged #phindinstant, aggregated by home.social.

  1. Introducing Phind-405B and faster, high quality #AI answers for everyone

    πŸš€ Phind-405B: New flagship #llm, based on Meta Llama 3.1 405B, designed for programming & technical tasks. #Phind405B

    ⚑ 128K tokens, 32K context window at launch, 92% on HumanEval, great for web app design. #Programming #AIModel

    πŸ’‘ Trained on 256 H100 GPUs with FP8 mixed precision, 40% memory reduction. #DeepSpeed #FP8

    ⚑ Phind Instant Model: Super fast, 350 tokens/sec, based on Meta Llama 3.1 8B. #PhindInstant

    πŸš€ Runs on NVIDIA TensorRT-LLM with flash decoding, fused CUDA kernels. #NVIDIA #GPUs

    πŸ” Faster Search: Prefetches results, saves up to 800ms latency, better embeddings. #FastSearch

    πŸ‘¨β€πŸ’» Goal: Help developers experiment faster, new features coming soon! #DevTools #Innovation

    phind.com/blog/introducing-phi