Skip to content
View ArtyomITA's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report ArtyomITA

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. llm2vec-8gb-offload-lossless-batching llm2vec-8gb-offload-lossless-batching Public

    Run LLM2Vec (Llama-3-8B) on an 8 GB GPU via RAM offload, and batch it without changing the embeddings (padding shifts them: cosine 0.76 vs 0.999956)

    Python

  2. vergilius vergilius Public

    Assistente AI locale per capire cosa succede nel mondo, ovunque: modelli piccoli, macchine modeste, nessun fornitore esterno. Odysseus + ShadowBroker.

    Python

  3. blackwell6000-qwen3.8-flash-next blackwell6000-qwen3.8-flash-next Public

    Blackwell 6000 + qwen 3.8 flash next: Swift 1.5 Qwen3.8 Flash-Next NVFP4 on one RTX PRO 6000 with vLLM, 4 users x 227k context, ~100 tok/s each, exact kernels + determinism fix

    Cuda

  4. flytollm flytollm Public

    The whole male Drosophila CNS connectome (166,700 neurons) used uncut as the spiking core of a language model. Scope, numbers, honest controls, pipeline animation.

    Python 1