🎯

Focusing

Sugato Ray sugatoray

🎯

Focusing

I am a Physicist turned Data Scientist + ML Practitioner. Research Interests: Data Science, ML, DL, Statistics, Math, Computing, LLMs.

75 followers · 176 following

View GitHub Profile

Recently created

Least recently created

Recently updated

Least recently updated

tomaarsen / export_locally.py

Created October 15, 2024 12:30

Export Sentence Transformer models to ONNX (+ optimization, quantization) & OpenVINO

	# requires sentence_transformers>=3.2.0
	from sentence_transformers import SentenceTransformer, export_optimized_onnx_model, export_dynamic_quantized_onnx_model

	# The model to export to ONNX (+ optimize, quantize), OpenVINO
	model_id = "mixedbread-ai/mxbai-embed-large-v1"
	# Where to save the exported models locally
	output_dir = model_id.replace("/", "-")

	onnx_model = SentenceTransformer(model_id, backend="onnx", model_kwargs={"export": True})
	onnx_model.save_pretrained(output_dir)

sugatoray / prompt.txt

Created October 6, 2024 14:53 — forked from philschmid/prompt.txt

	Begin by enclosing all thoughts within <thinking> tags, exploring multiple angles and approaches.
	Break down the solution into clear steps within <step> tags. Start with a 20-step budget, requesting more for complex problems if needed.
	Use <count> tags after each step to show the remaining budget. Stop when reaching 0.
	Continuously adjust your reasoning based on intermediate results and reflections, adapting your strategy as you progress.
	Regularly evaluate progress using <reflection> tags. Be critical and honest about your reasoning process.
	Assign a quality score between 0.0 and 1.0 using <reward> tags after each reflection. Use this to guide your approach:

	0.8+: Continue current approach
	0.5-0.7: Consider minor adjustments
	Below 0.5: Seriously consider backtracking and trying a different approach

philschmid / prompt.txt

Last active November 15, 2024 06:01

	Begin by enclosing all thoughts within <thinking> tags, exploring multiple angles and approaches.
	Break down the solution into clear steps within <step> tags. Start with a 20-step budget, requesting more for complex problems if needed.
	Use <count> tags after each step to show the remaining budget. Stop when reaching 0.
	Continuously adjust your reasoning based on intermediate results and reflections, adapting your strategy as you progress.
	Regularly evaluate progress using <reflection> tags. Be critical and honest about your reasoning process.
	Assign a quality score between 0.0 and 1.0 using <reward> tags after each reflection. Use this to guide your approach:

	0.8+: Continue current approach
	0.5-0.7: Consider minor adjustments
	Below 0.5: Seriously consider backtracking and trying a different approach

sugatoray / pipeline_parallel.py

Created October 2, 2024 17:20 — forked from 3outeille/pipeline_parallel.py

Self contained example of how pipeline parallel works (AFAB and 1F1B) in 200 LOC

	#VERBOSE=0 torchrun --nproc_per_node 3 self_contained_pp_LOC.py
	import os, random, numpy as np, torch, torch.nn as nn, torch.distributed as dist, torch.nn.functional as F
	from torch.optim import AdamW
	from torch.utils.data import DataLoader, DistributedSampler
	from datasets import load_dataset
	from transformers import AutoConfig, AutoModelForCausalLM, AutoTokenizer

	STEP, local_rank, world_size, verbose = 0, int(os.environ["LOCAL_RANK"]), int(os.environ["WORLD_SIZE"]), os.environ.get("VERBOSE", "0") == "1"

	def set_all_seed(seed):

3outeille / pipeline_parallel.py

Last active November 15, 2024 19:37

Self contained example of how pipeline parallel works (AFAB and 1F1B) in 200 LOC

	#VERBOSE=0 torchrun --nproc_per_node 3 self_contained_pp_LOC.py
	import os, random, numpy as np, torch, torch.nn as nn, torch.distributed as dist, torch.nn.functional as F
	from torch.optim import AdamW
	from torch.utils.data import DataLoader, DistributedSampler
	from datasets import load_dataset
	from transformers import AutoConfig, AutoModelForCausalLM, AutoTokenizer

	STEP, local_rank, world_size, verbose = 0, int(os.environ["LOCAL_RANK"]), int(os.environ["WORLD_SIZE"]), os.environ.get("VERBOSE", "0") == "1"

	def set_all_seed(seed):

hanxiao / testRegex.js

Last active November 14, 2024 15:40

Regex for chunking by using all semantic cues

	// Updated: Aug. 20, 2024
	// Run: node testRegex.js whatever.txt
	// Live demo: https://jina.ai/tokenizer
	// LICENSE: Apache-2.0 (https://www.apache.org/licenses/LICENSE-2.0)
	// COPYRIGHT: Jina AI
	const fs = require('fs');
	const util = require('util');

	// Define variables for magic numbers
	const MAX_HEADING_LENGTH = 7;

awni / mflux_steps.md

Last active November 15, 2024 17:40

Setup the repo

git clone [email protected]:filipstrand/mflux.git
cd mflux && pip install -r requirements.txt

Make a run script

Name this anything, maybe flux.py. Make sure to update the two paths marked below.

awni / l3min.py

Last active November 2, 2024 16:06

A minimal, fast implementation of Llama 3.1 in MLX.

	"""
	A minimal, fast example generating text with Llama 3.1 in MLX.

	To run, install the requirements:

	pip install -U mlx transformers fire

	Then generate text with:

	python l3min.py "How tall is K2?"

sugatoray / demo-whisper-medusa.ipynb

Created August 6, 2024 05:16

demo-whisper-medusa

Sorry, something went wrong. Reload?

Sorry, we cannot display this file.

Sorry, this file is invalid so it cannot be displayed.

awni / mlx_lm_openai.md

Last active August 16, 2024 05:14

MLX LM with the OpenAI Python Package

1. Install

Install MLX LM and openai:

pip install mlx-lm openai

NewerOlder