Interesting links, 20/04/2026
Claude summary/sorting of old tabs
- Tabs
- Irish Phonology & Texts
- G2P & Phoneme Processing
- Speech Alignment & ASR
- Speech Synthesis & TTS
- Audio & Speech Models
- Swedish & Scandinavian Language Resources
- Hungarian Language Learning
- Sami, Kven & Nordic Minority Languages
- Transliteration & Romanization
- NLP / LLM Tools
- Computer Vision & Multimodal
- Robotics & Motion Capture
- ML Frameworks & Training Tools
- LibriVox & Digital Libraries
- Entertainment & Media
- Retro Computing
- Misc
Tabs
Irish Phonology & Texts
Wikisource: A Dialect of Donegal (Quiggin)
- A Dialect of Donegal/Introductory
- A Dialect of Donegal/The Vowel System
- A Dialect of Donegal/Texts/Áindrías an Ime
- A Dialect of Donegal/Texts/Eóin Ua Míodhchán agus an Sionnach
- A Dialect of Donegal/Texts/Éamonn Ua Ciórrthais
- Page:Quiggin Dialect of Donegal 0004.png
- Page:Quiggin Dialect of Donegal 0026.png
- Page:Quiggin Dialect of Donegal 0198.png
- Page:Quiggin Dialect of Donegal 0199.png
Wikisource: Die araner mundart (Finck)
- Die araner mundart/Wort-Register
- Seite:Die araner mundart.djvu/29
- Seite:Die araner mundart.djvu/288
- Seite:Die araner mundart.djvu/493
Wikisource: Kerry Irish
Wikisource: Desi-Irish Phonology
- Page: A contribution to the phonology of Desi-Irish …/17
- Page: A contribution to the phonology of Desi-Irish …/42
- Page: A contribution to the phonology of Desi-Irish …/76
Wikisource: Other Irish
- Page:Simple Lessons in Irish, Part 1 - O’Growney.pdf/14
- Wikisource Community Collaboration/Monthly Challenge/January 2025
Other Irish Resources
- Stór Scéalta - Leigh Leat
- Abair Liom, An Rud Maith é an Drón? - Leigh Leat
- Bríd na nAmhrán - Leigh Leat
- Leabhar mhí na Bealtaine 2018 - ClubLeabhar.com
- SpeakGaelic on Instagram
G2P & Phoneme Processing
CharsiuG2P
- charsiu/g2p_multilingual_byT5_small_100 · Hugging Face
- charsiu/g2p_multilingual_byT5_small
- charsiu/IPATokenizer tokenizer_config.json
- charsiu/tokenizer_en_cmu
- charsiu models (HuggingFace, p.0)
- charsiu models (HuggingFace, p.1)
- lingjzhu/CharsiuG2P
- CharsiuG2P/src/g2p.py
- CharsiuG2P/src/model.py
- CharsiuG2P/src/data_utils.py
- CharsiuG2P/src/clean_wordlist.py
- CharsiuG2P/src/clean_dicts.py
- CharsiuG2P/src/ByT5_MoE.py
- CharsiuG2P/src/train.py
- CharsiuG2P/src/CharsiuG2P.py
- lingjzhu (jzhu)
- ByT5-Finetuning-Datasets.ipynb (Colab)
- charsiu_tutorial.ipynb (Colab)
CLAP-IPA
- The taste of IPA: Towards open-vocabulary keyword spotting and forced alignment in any language
- lingjzhu/clap-ipa
- Clap: Complete Guide 2024
Other G2P Tools
- rhasspy/gruut: tokenizer, text cleaner, and phonemizer
- g2p · PyPI
- tabahi/contexless-phonemes-CUPE
- Tabahi/CUPE-2i example.py
- tabahi/bournemouth-forced-aligner
- pavelsof/ipavec: IPA alignment using vector representations
- panphon/panphon/data
- panphon/ipa_bases.csv
- panphon/ipa_all.csv
- Generating Phonological Feature Vectors with SoundVectors and CLTS
- wikipron/extract/lat.py
- wikipron/extract/jpn.py
- ishine/PnG-BERT
- g2p-calculate-per.ipynb (Colab)
- lrec-g2p-pivot (Overleaf)
- Anonymous GitHub
Speech Alignment & ASR
Kaldi
- kaldi.alignment — PyKaldi 0.1.1 documentation
- kaldi-align/kaldi_align
- kaldi-align/align2json.py
- sv_kaldi-montreal/acoustic_model/path.sh
Other Aligners
- PocketSphinx Documentation
- xinjli/allosaurus: universal phone recognizer for 2000+ languages
- allosaurus/app.py
- NeMo Forced Aligner (NFA)
- lingjzhu/charsiu: neural phonetic aligner
- charsiu_demo.ipynb (Colab)
- alopez/hmmalign: HMM word aligner for SMT
- dreamt/aligner/align
- Cross Attention with Monotonic Alignment for Speech Transformer
ASR Models & Corpora
- lingjzhu/zipa: efficient speech models for multilingual phone recognition
- zipa/zipformer_crctc/ctc_decode.py
- Open Whisper-style Speech Models (OWSM) collection
- espnet/owsm_v3.2
- espnet/owsm_v3.1_ebf_small_lowrestriction
- pyf98/owsm_ctc_v3.1_1B
- icefall/egs/speech_llm/ASR_LLM
- 04-lhotse-shar.ipynb (Colab)
- huggingface_wav2vec/Dockerfile
- Pooya-Fallah/whisper-tiny-finetune
- jimbozhang/speechocean762
- mispeech/speechocean762 dataset
- RNN-Transducer with stateless prediction network (Google Research)
- MMSpeech: Multi-modal Multi-task Encoder-Decoder Pre-training
- QwenLM/Qwen2-Audio
- qwen2-audio-finetune notebook
- Qwen/Qwen2-Audio-7B-Instruct
- Qwen2-Audio blog post
- versa/bin/espnet_scorer.py
- shinjiwlab
- On the Effects of Heterogeneous Data Sources on Speech-to-Text Foundation Models
- An Illustrated Tour of Applying BERT to Speech Data
- SpeechBERT: Audio-and-text Jointly Learned Language Model
- Edge-Punct-Casing
- A light-weight punctuation and word casing model for on-device streaming ASR
Speech Synthesis & TTS
- microsoft/VibeVoice-1.5B
- microsoft/VibeVoice: Frontier Open-Source Text-to-Speech
- microsoft/SpeechT5
- Fine-tuning SpeechT5 - HuggingFace Audio Course
- HoseinAzad/SpeechT5-Non-English-TTS
- SpeechT5 TTS Fine-tuning.ipynb (Colab)
- hfblog/speecht5.md
- 18. Fine-tuned SpeechT5 (Kaggle)
- rhasspy/piper: fast, local neural TTS
- piper/TRAINING.md
- piper-phonemize/phoneme_ids.cpp
- piper-voices hu_HU-berta-medium.onnx.json
- Kokoro TTS (HuggingFace Space)
- Zonos v0.1 beta release
- Zonos/sample.py
- LLaSA: Large Language and Structured Data Assistant
- Llasa: Scaling Train-Time and Inference-Time Compute for Speech Synthesis
- LLaSA_training/run_slurm.sh
- EzAudio: Enhancing Text-to-Audio Generation
- SoundCTM
- sony/soundctm
- Sony/soundctm teacher ckpt
- VoiceRestore (HuggingFace Space)
- skirdey/voicerestore
- dioco-group/jenny-tts-dataset
- tihu-nlp/tihu: Persian Text-To-Speech
- karim23657/Persian-tts-coqui
- ttsds benchmark
- TTS Arena (HuggingFace Space)
- Speech Synthesis Workshop 2025 - University of Groningen
- North Sami Text-to-Speech (GiellAlt)
- North Sami TTS - recording the voices
- North Sami TTS - IPA generating pipeline
- speech-sme/CaseInNumerals.md
- Hállansyntesa - Divvun
- Text-to-speech - Divvun
- Borealium voice (se female)
- TTSTextNormalization/converters
- tarepan/SpeechMOS
Audio & Speech Models
Audio Embeddings & CLAP
- microsoft/CLAP: Learning audio concepts from natural language
- CLAP/models/audio.py
- CLAP/examples/zero_shot_predictions.py
- CLAP/examples/esc50_dataset.py
- qiuqiangkong/audioset_tagging_cnn
- bakhtos/PANNs
- marl/openl3: Open-source deep audio and image embeddings
- audio_embeddings/Audio Search notebook
- Universal Speech Model (Google)
- facebookresearch/audiobox-aesthetics
Textless / Self-supervised Speech
- facebookresearch/textlesslib issues
- facebookresearch/speech-resynthesis
- Berkeley-Speech-Group/sylber: Syllabic Embedding Representation from Raw Audio
- [2202.09729] It’s Raw! Audio Generation with State-Space Models
- [2403.05010] RFWave: Multi-band Rectified Flow for Audio Waveform Reconstruction
- pYAAPT — AMFM_decompy documentation
- Source–filter model - Wikipedia
Unsupervised Speech Segmentation
- kamperh/vqwordseg
- kamperh/VectorQuantizedCPC
- bshall/VectorQuantizedCPC
- kamperh/segmentalist
- kamperh/bucktsong_segmentalist
- kamperh/dpdp_aernn
- kamperh/speech_dtw
- kamperh/speech_correspondence/train_stacked_dae.py
- kamperh/bayes_gmm/fbgmm.py
- kamperh/lecture_dtw_notebook
- melsner/neural-segmentation
- ZRTools/srailsdisc/srails_disc.c
- felixkreuk/UnsupSeg: Self-Supervised Contrastive Learning for Unsupervised Phoneme Segmentation
Corpora & Benchmarks
- AusTalk (web archive)
- Buckeye Corpus
- Buckeye Corpus publications
- Buckeye SpeechSearcher Manual
- scjs/buckeye Python library
- wav2gloss (CMU LTI)
- wav2gloss/cocoon-glosses dataset
- wav2gloss/NINJAL-Ainu-Folklore
- wav2gloss/odin
- datashare.ed.ac.uk 10283/3443 README
- RESOURCEFUL-2025
Speech Phonetics Tools
- espsfree/sgram.1
- The Berkeley Phonetics Machine (ISCA paper)
- rsprouse (Ronald L. Sprouse)
- Anna Lewington Collection of Matsigenka and Yine Recordings
- IFC formant tracker (Berkeley)
- ucblingmisc/perl/convertlabel
- klsyn/c/klsyn.c
- audiolabel/audiolabel.py
- VGG Image Annotator
- Linear Prediction Analysis (Theory) - Amrita Virtual Lab
- EffectsExplained - sox_ng
- SpeechSearcher Manual
- Microsoft Word - transcription-ELL.pdf
Swedish & Scandinavian Language Resources
Braxen / Swedish Pronunciation
- sprakbankental/braxen commit patch
- Braatöy Trygve - Bokbörsen
- 2025.nodalida-1.71.pdf
- FULLTEXT01.pdf (DiVA)
- BSF_7E302.pdf
Swedish Dialects
- SweDia 2000 publications
- www.swedia.nu
- SweDia Snabbmeny
- SweDia: Äldre kvinna, Össjö
- Swedish accent navigation - Lunds universitet
- Aspects of North Swedish intonational phonology - Lunds universitet
- 1624497.pdf (Swedish accent)
- The traditional Swedish dialect areas (ResearchGate figure)
- Älvdalska - Älvdalens kommun
- Nordic Dialect Corpus
- Nordic Dialect Corpus: Transcription
- Transkripsjonsrettleiing for ScanDiaSyn
- User Manual for Nordic Dialect Corpus
- LIA Norwegian
- textlab/spoken_norwegian_resources
Swedish NLP
- Folkets lexikon
- Swedish phonology - Wikipedia
- Dependensparsningsmodell: Stanza - Språkbanken Text
- UD_Swedish-LinES
- UD_Swedish-PUD
- UniversalDependencies/UD_Swedish-PUD
- robertostling/efselab: Efficient Sequence Labeling
- UD_Slovenian-SST/sl_sst-ud-test.conllu
- Universal Dependencies for Swedish Sign Language
- Modul:uttal - Swedish Wiktionary
- language-resources/festus/alignables-util.cc (Google)
Danish
- Danish phonology - Wikipedia
- Help:IPA/Danish - Wikipedia
- Category:Danish terms with IPA pronunciation - Wiktionary
- abebrødtræ - Wiktionary (Danish)
- afsender - Wiktionary (Danish)
Hungarian Language Learning
YouTube
- Barátok közt - Wikipedia
- MagyarOK - MI VAN A TÁSKÁDBAN?
- A visit (IMP + COND) – A2-B1 - Hungarian for foreigners
- Robinson Crusoe története 1/20 #a2level #b1level
- HUNGARIAN Conversation 1 ⭐ (A2-B1)
- Hungarian REACTIONS When You’re Surprised - Slow Hungarian Dialogues
- MagyarOK - Régen találkoztunk már!
- Ep 137: When your partner doesn’t speak your mother tongue [B1]
- Ep 122: 5+1 Painful Truth about Language Learning [B1/B2]
- Ep 99: How Hungarians think and talk about money [A2/B1]
- In the buffet + What is (s)he wearing? A1-B1
- Hungarian Dialogue: Friends Making Plans [A2-B1]
- MagyarOK - BEMUTATKOZÁS 1
Grammar References
Learning Resources
- The Best Resources for Learning Hungarian • Catch Budapest
- FREE Course: The Fast Lane to Understand Hungarian • Catch Budapest
- Hungarian Language Resources
- Stream Radio in Hungarian
- Hungarian learning with free materials - Hungarian Kati
- Free Hungarian E-book - Kati és Mari - Hungarian Kati
- Best traditional textbook for Hungarian : r/hungarian
- Learning resources - Let’s Learn Hungarian!
- Let’s Learn Hungarian! Podcast - Apple Podcasts
- Free Hungarian Learning Materials - Talkpal
- FSI Hungarian
- FSI - Hungarian Basic Course - Volume 1
- Learn Hungarian online - Loecsen
- Smart Hungarian Short Stories Course • Catch Budapest
- Graded readers for Hungarian? - Language Learning Stack Exchange
- Language Professor UNLOCKS 100 year old Learning Method - YouTube
Magyar Elektronikus Könyvtár (MEK) Audiobooks
- Hungarian Electronic Library - Advanced search
- MEK audio search (MP3)
- MEK 02286 (audio)
- MEK 02286 MP3 index
- MEK 25889
- MEK 25889 MP3 index
- MEK 13438: Koldus és királyfi (Mark Twain)
- Koldus és királyfi OCR PDF
- MARK TWAIN: KOLDUS ÉS KIRÁLYFI (full text)
- Katalógus - Magyar Vakok és Gyengénlátók hangoskönyvtár
- BLAHALOUISIANA - veszélyes utcák (YouTube)
Sami, Kven & Nordic Minority Languages
- Kven language - Wikipedia
- STR-T – Äänitorvi kotimaisele kieleliselle ja kulttuuriselle alkukansale
- NRK Kveeni
- Kampen for å bevare kvensk kultur og språk – NRK
- Kvensk språkdag - Lær fem kvenske ord og fraser – NRK
- NRK Kvääni - Wikipedia (no)
- Halti kvenkultursenter IKS
- Varanger museum
- Kvensk-norsk-kvensk ordbok – Kainun Institutti
- Nybegynnerkurs kvensk, 10 deler – Kainun Institutti
- Kvensk.no - Læringsressurser til kvenskundervisning
- 10. kapitteli - Kvensk
- Nettidigisanat - Neahttadigisánit
- Nettidigisanat - Neahttadigisánit (vep/fin)
- Nettidigisanat - Neahttadigisánit (fit/swe)
- oahpaa.no (North Sami)
- Divvun - sámi giellateknologiija
- čuođi — Wiktionnaire
Transliteration & Romanization
- cldr/common/transforms (unicode-org)
- CLDR: Russian-Latin-BGN.xml
- Romanization of Russian - Wikipedia
- Romanization of Arabic - Wikipedia
- Romanization of Greek - Wikipedia
- Wikipedia:Naming conventions (Greek)
- Wikipedia:Naming conventions (Cyrillic)
- Montenegrin alphabet - Wikipedia
- Wikipédia:Romanização/Russo (pt)
- Wikipédia:Romanização/Grego (pt)
- Wikipédia:Romanisation du russe (fr)
- Transcription du russe en français
- Romanisation du grec (fr)
- Aiuto:Greco moderno (it)
- Aiuto:Greco antico (it)
- Wikipedia:Transliteratie- en transcriptiegids (nl)
- Wikipedia:Transliteratie- en transcriptiegids/Russisch (nl)
- Wikipedia:Transliteratie- en transcriptiegids/Latijn en Grieks (nl)
- Wikipedia:Transliteración y transcripción (es)
- Transkribering av östslaviska språk (sv)
- Transkribering av ukrainska (sv)
- Kategori:Transkriptionssystem (sv)
- Direkt system för translitterering av bulgariska (sv)
- A Multitask Learning Approach for Diacritic Restoration (ACL 2020)
NLP / LLM Tools
Search & RAG
- jina-ai/node-DeepResearch
- Jina Reader API
- Open-source DeepResearch – HuggingFace
- smolagents/open_deep_research/text_inspector_tool.py
- stanford-oval/WikiChat
- stanford-oval/wikipedia_20240801 bge-m3 qdrant index
- BAAI/bge-m3
- rag-langchain-audio-data/main.py
- Retrieval Augmented Generation on audio data with LangChain
- Azure-Samples/aisearch-openai-rag-audio
- Tutorial: ChatGPT Over Your Data (LangChain)
- facebook/rag-token-nq
- asg017/sqlite-vec
- asg017/sqlite-ecosystem
- asg017/sqlite-lembed
- ACL Anthology
- Findings of EACL 2024
LLM Training & Fine-tuning
- OpenRLHF/OpenRLHF: Scalable RLHF Framework
- GRPO Trainer (HuggingFace TRL)
- 7B Model and 8K Examples: Emerging Reasoning with RL (simplerl-reason)
- The Hundred-Page Language Models Book (Burkov)
- aburkov/theLMbook
- The Hundred-Page Machine Learning Book
- aburkov/theMLbook
- huggingface/alignment-handbook
- martyn/safetensors-merge-supermario
- facebookresearch/lingua: Meta Lingua
- lingua/apps/fastRNN/hawk/core_hawk.py
- facebookresearch/LayerSkip (ACL 2024)
- pytorch/torchtitan
- pytorch/torchtune
- Scaling FineWeb to 1000+ languages (HuggingFace Space)
- Contributing to multilingual evaluations · huggingface/lighteval
Models
- CohereForAI/aya-101
- google/mt5-xxl
- ai21labs/Jamba-v0.1
- configuration_jamba.py
- microsoft/Phi-3-mini-4k-instruct sample_finetune.py
- Fine-tune Phi-3 for sentiment analysis (Kaggle)
- AI-MO/NuminaMath-7B-TIR
- DeepSeek-Math LICENSE-MODEL
- Phi-4 (HuggingFace Space)
- HuggingFaceTB/SmolVLM-500M-Instruct
- neonbjb (James Betker)
- ocotillo/model_loader.py
- QwenLM/Qwen2-Audio
- Qwen/Qwen2-Audio-7B-Instruct
- facebook/opt-iml-1.3b
Evaluation & Interpretability
- Prometheus-Eval
- Prometheus 2
- Gemma Scope (Google DeepMind)
- google-deepmind/mishax
- openai/sparse_autoencoder
- SAE viewer (GPT2-small)
- dottxt-ai/outlines: Structured Text Generation
- LoRAX + Outlines: Better JSON Extraction
- MMBench
Entity Recognition & Graph
- urchade/GLiNER (NAACL 2024)
- urchade/gliner_multi-v2.1
- medieval-data/gliner_multi-v2.1-medieval-latin
- SapienzaNLP/relik (ACL 2024)
- Entity Linking and Relationship Extraction With Relik in LlamaIndex
- eustlb/speech-to-speech
- huggingface/dataspeech requirements.txt
Parsing & Disfluency
- pariajm/english-fisher-annotations
- pariajm/joint-disfluency-detector-and-parser
- nikitakit/self-attentive-parser
- harvardnlp/pytorch-struct
Computer Vision & Multimodal
Vision Language Models
- microsoft/Florence-2-large-ft
- Florence-2 sample_inference.ipynb
- Fine-tuning Florence-2 (HuggingFace blog)
- andimarafioti/florence2-finetuning
- microsoft/Phi-3-vision-128k-instruct
- Microsoft Phi-3-Vision-128k (HuggingFace Space)
- microsoft/GLIP: Grounded Language-Image Pre-training
- microsoft/OmniParser
- microsoft/OmniParser (HuggingFace)
- Cambrian-1: Vision-Centric Multimodal LLMs
- cambrian-mllm/cambrian
- nyu-visionx/cambrian-34b
- AIGText/Glyph-ByT5
- DAMO-NLP-SG/multimodal_textbook dataset
- DAMO-NLP-SG/multimodal_textbook repo
- OFA-Sys/OFA (ICML 2022)
- OFA-Sys/ofa-huge
- salesforce/LAVIS (xgen-mm)
- Paper: Ola - Omni-Modal Language Model
- vidore/colpali
- tonywu71/vidore-benchmark
- colpali/clip_baselines.py
- InferSent/models.py
Human Pose & Gestures
- facebookresearch/sapiens: High-resolution models for human tasks
- facebook/sapiens (HuggingFace)
- Sapiens Segmentation (HuggingFace Space)
- MotionBERT: Unified Human Motion Representations
- MotionBERT/lib/utils/vismo.py
- 3D Human Pose Estimation
- Human3.6M Benchmark
- VideoPose3D with Detectron2
- [2408.00370] DiM-Gesture: Co-Speech Gesture Generation
- zf223669/DiMGestures
- zf223669/DiffGesture (CVPR 2023)
- Learning Hierarchical Cross-Modal Association for Co-Speech Gesture
- Audio-Driven Co-Speech Gesture Video Generation
- alvinliu0/ANGIE (NeurIPS 2022)
- zf223669/ListenDenoiseAction (SIGGRAPH 2023)
- [2407.21491] Generative Expressive Conversational Speech Synthesis
- AI-S2-Lab/GPT-Talker
- Paper: MotionLab - Unified Human Motion Generation
- omnihuman-lab.github.io
- STAR: A Benchmark for Situated Reasoning in Real-World Videos
- NVlabs/LSM: Large Spatial Model (NeurIPS 2024)
- Large Spatial Model project page
- Cross-Dialect Text-to-Speech In Pitch-Accent Language (IEEE)
Image/3D Generation
- jy0205/Pyramid-Flow
- Allenai Objaverse
- allenai/objaverse-xl dataset
- Grokking Diffusion Models
- leejet/stable-diffusion.cpp
- axodox/axodox-machinelearning (C++ ONNX)
Robotics & Motion Capture
ROS & MoveIt
- MoveIt Motion Planning Framework
- Robots - MoveIt
- MoveIt 2 Documentation
- Kinematics Cost Functions — MoveIt
- Planning Around Objects — MoveIt
- Move Group Python Interface (kinetic)
- pal-robotics/tiago_tutorials
- Robots/TIAGo/Tutorials - ROS Wiki
- Robots/TIAGo/Tutorials/MoveIt/Planning_joint_space
- rviz/DisplayTypes/TF
- ari_robot/ari_description/robots/ari.urdf.xacro
- b-adkins/thor_arm_moveit
- THÖR DATA SET
- Awesome Robotics Libraries
- Ly0n/awesome-robotic-tooling
- Unity-Technologies/Unity-Robotics-Hub
- Gazebo features
Motion Capture / BVH
- matt-graham/bvh-tools
- bvhtoolbox · PyPI
- bvhsdk/egocentriccoord.py
- HKUST-HCI/bvh_broadcaster
- bvh_broadcaster/cmu_mocap_bvh.md
- pose-prediction/expmap.py
- pytorch3d.transforms documentation
LeRobot
- huggingface/lerobot
- lerobot (LeRobot HuggingFace)
- lerobot/xarm_push_medium dataset
- lerobot/aloha_static_pro_pencil dataset
- lerobot/diffusion_pusht
- huggingface/gym-pusht
- google-research/rlds
- google-deepmind/dm_env
- google-deepmind/mujoco_menagerie
- robot-descriptions/robot_descriptions.py
- robot-descriptions/awesome-robot-descriptions
ML Frameworks & Training Tools
PyTorch Ecosystem
- pytorch/torchtitan
- torchtitan/train_configs
- pytorch/torchtune
- pytorch/ao: quantization and sparsity
- pytorch/botorch: Bayesian optimization
- facebook/Ax: Adaptive Experimentation Platform
- GPyTorch
- pytorch/torchcodec
- pytorch/extension-cpp
- pytorch/translate
- pytorch/QNNPACK
- pytorch/FBGEMM/fbgemm_gpu/src
- pytorch/data/examples/audio/librispeech.py
- Multi-node-training on slurm with PyTorch (gist)
- facebookresearch/moco/main_lincls.py
- vision/references/classification/train.py
- Supported NumPy features — Numba
- numpy.lib.stride_tricks.as_strided
- LoRAX + Outlines JSON extraction
- JAX-Toolbox/rosetta/projects/t5x
- Scaling up linear programming with PDLP (Google)
ONNX & Model Conversion
- Convert Keras models to ONNX (Medium)
- onnx2pytorch · PyPI
- ENOT-AutoDL/onnx2torch
- onnxruntime-wav2vec/wav2vec2onnx.ipynb
- huggingface/candle: Minimalist ML framework for Rust
Misc ML
- LoGAH: Predicting 774M-Parameter Transformers
- facebookresearch/ppuda (NeurIPS 2021)
- LoGAH/dataset_generator/vit_generator.py
- KAN-GPT-2/toy_functions.py
- SynodicMonth/ChebyKAN
- facebookresearch/ppuda
- ring-attention/flash_attn_triton.py
- How I Studied LLMs in Two Weeks
- hesamsheikh/ml-retreat
- Integer addition algorithm: 95% energy reduction for AI
- Replicate — Run AI with an API
LibriVox & Digital Libraries
- LibriVox search
- LibriVox: The Partition of Europe (JSON)
- LibriVox Community Podcast #153
- LibriVox: Short Story Collection Vol. 081
- LibriVox: David Copperfield Band 08 (Dickens)
- Project Gutenberg: Russian Fairy Tales (Ralston)
- Index:Ernest Hemingway - A Farewell to Arms.pdf (Wikisource)
- Wikimedia Commons: English pronunciation category
- Wikimedia Commons: Voice of America pronunciation
Entertainment & Media
- Prime Video: The Perks of Being a Wallflower
- play.hbomax.com
- Netflix
- Kongen Befaler S12 E03 (with English subs) : r/panelshow
- Kongen Befaler S12 E03.mp4 - Google Drive
- Kongen Befaler S11 E06.mp4 - Google Drive
- Pentulive 24/7 - Yle Areena
- BBC Sounds - I’m Sorry I Haven’t A Clue
- 5 Signs You’re A High-Masking Autistic With ADHD - YouTube
- How To Spot Autism in High-Masking Women and Girls - YouTube
Retro Computing
- Commodore BASIC as a Scripting Language for UNIX and Windows – pagetable.com
- mist64/cbmbasic
- Create your own Version of Microsoft BASIC for 6502 – pagetable.com
- Apple II: Making an Apple II disk for an emulator · cc65/wiki
- Apple ProDOS 8 development · GitHub
- Applesoft Lite - Applesoft BASIC for the Replica-1
- txgx42/applesoft-lite
- JOYCE for UNIX (PCW emulator)
- PCW PD Catalogue
- LocoScript 1 file format
Misc
- Anonymous GitHub
- Yaak API client - now Open Source
- yaakapp/app
- Aalto University: AScI International Summer Research Programme
- Research:Newsletter/2024/May - Wikimedia Meta
- Export all messages from a Signal conversation? : r/signal
- What is the location of WhatsApp’s encryption key file? : r/DataHoarder
- Extracting Messages from Signal Desktop
- ElDavoo/wa-crypt-tools
- B16f00t/whapa: WhatsApp Parser Toolset
- IBM: Speech recognition history
- RESEARCH TEAM AT I.B.M. DEVELOPS A NEW COMPUTER - NYT 1984
- Redox OS: Rust-Based Alternative to Linux
- Orbital Desktop Environment
- Category:Sign languages - Wikimedia Commons
- Category:SignWriting by language - Wikimedia Commons
- Category:Polish Sign Language in SignWriting
- alphacep/awesome-speech
- KTH speech publications (1846.pdf)
- KTH speech publications (50år_8.pdf)
- Efficient data generation for source-grounded dialogs (Google Research)
- google-research-datasets/MISeD
- Yale-LILY/QMSum
- moby/moby: The Moby Project
- moby/datakit
- google-deepmind/natural-plan
- google-deepmind/open_spiel
- Pronunciation Adaption at the Lexical Level (RU repository)