This is your work, valued
panflute. An Pythonic alternative to John MacFarlane's pandocfilters, with extra helper functions
★ 555reghdfe. Linear, IV and GMM Regressions With Any Number of Fixed Effects
★ 250ftools. Fast Stata commands for large datasets
★ 149ivreghdfe. Run IV/2SLS with many levels of fixed effects (i.e. ivreg2+reghdfe)
★ 112ppmlhdfe. Poisson pseudo-likelihood regression with multiple levels of fixed effects
★ 73panflute-filters. Pandoc filters that use Panflute
★ 62clv-locro. Wrapper for Chromium screen-ai OCR
★ 36quipucamayoc. dev repo for article
★ 33stata-require. Enforce exact/minimum versions of community-contributed packages.
★ 19stata-misc. Miscellaneous Stata Commands
★ 13pandocmk. Python library that simplifies running Pandoc
★ 11parse-smcl. Parse SMCL Help Files into Markdown and HTML
★ 10sublime-stata. Sublime Text package for Stata (improved syntax, snippets, and shortcuts)
★ 9data-covid-minsa. limpieza rapida de datos de covid de minsa
★ 7stata-setroot. Find the root path of a project and set it as a global variable
★ 7useful-stata. Do file that quickly installs useful Stata programs
★ 7stata-schemes. Some personal Stata figure schemes (styles)
★ 6fedplot. R package to create ggplot2 charts in the style used by the Fed's FSR
★ 5egenfast. Because -egen- is too slow
★ 4StataEditor. Stata Editor for Sublime Text 3
★ 4stata-ensemble-ocr. Stata package that combines different versions of the same variable, each obtained from different OCR engines or scans
★ 4stata-economics. Economics Lesson with Stata
★ 3camelot-1. A Python library to extract tabular data from PDFs
★ 3xtsmooth. Panel version of Stata's -smooth- command
★ 3camelot. Camelot: PDF Table Extraction for Humans
★ 2Statapackagesearch. Search for (missing) packages in Stata code
★ 2quipu. Manage Stata Estimation Results
★ 2bcrpuse. Stata module to Import data from the Peruvian Central Bank (BCRP)
★ 2arxivate. Python package that prepares a tex project for arxiv submission
★ 2python-frontmatter. Parse and manage posts with YAML frontmatter
★ 1stata-utilities. Miscellaneous Stata utilities
★ 1MarkdownEditing. Powerful Markdown package for Sublime Text with better syntax understanding and good color schemes.
★ 1binscatter. Stata module to generate binned scatterplots —
★ 1geocode_ip. Geocode IP addresses in Stata
★ 1panflute-dockerfiles. Experimental docker files for panflute
★ 1panzer. pandoc + styles
★ 1estout. Stata module to make regression tables
★ 1xhdfe-xfe. Linear regression with multiple high-dimensional fixed effects, in Stata, Python and R on one C++ core (reghdfe-comparable; optional CUDA)
★ 7clv-locro. Wrapper for Chromium screen-ai OCR
★ 40arxivate. Python package that prepares a tex project for arxiv submission
★ 2stata-ctools. C-accelerated drop-in replacements for Stata programs
★ 4Doxa. A Local Adaptive Thresholding framework for image binarization written in C++, with JS, Python and MATLAB bindings. Implementing: Otsu, Bernsen, Niblack, Bradley, Sauvola, Wolf, Gatos, NICK, Su, T.R. Singh, WAN, ISauvola, Feng, Phansalkar, AdOtsu, along with DRDM and all Pseudo Merics.
★ 194fasthtml. The fastest way to create an HTML app
★ 7kadodown. Tools for building Stata package documentation websites
★ 7talk2arxiv. Talk to any ArXiv paper using ChatGPT
★ 529lancedb. Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
★ 11kdid2s_stata. Two-Stage Difference-in-Differences following Gardner (2021)
★ 35latex-templates. latex templates
★ 290stata-require. Enforce exact/minimum versions of community-contributed packages.
★ 19gginnards. R package extending 'ggplot2' with manipulation and debugging tools.
★ 28fedplot. R package to create ggplot2 charts in the style used by the Fed's FSR
★ 5efficient-geopandas-workshop. Workshop materials for the Writing an efficient code for GeoPandas and Shapely in 2023
★ 88cli-compat-stata. cli-compat-stata
★ 3johnjosephhorton.github.io. HTML
★ 4OCRmyPDF. OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
★ 34ktidytable. Tidy interface to 'data.table'
★ 474DewarpNet. Code for the paper "DewarpNet: Single-Image Document Unwarping With Stacked 3D and 2D Regression Networks" (ICCV '19)
★ 624algpseudocodex. LaTeX package for typesetting pseudocode.
★ 53quipucamayoc. dev repo for article
★ 33psacalc_supports_reghdfe. A minor revision to `psacalc` to add support for `reghdfe`
★ 23khroma. Colour Schemes for Scientific Data Visualization - :exclamation: Moved to https://codeberg.org/tesselle/khroma
★ 217MyST-Parser. An extended commonmark compliant parser, with bridges to docutils/sphinx
★ 882sphinx-autobuild. Watch a Sphinx directory and rebuild the documentation when a change is detected. Also includes a hot-reload web server.
★ 608mkdocstrings. :blue_book: Automatic documentation from sources, for MkDocs.
★ 2.1kunbuch. Compile markdown into an html and pdf book based on pandoc.
★ 204DocTr. The official code for “DocTr: Document Image Transformer for Geometric Unwarping and Illumination Correction”, ACM MM, Oral Paper, 2021.
★ 438xtsmooth. Panel version of Stata's -smooth- command
★ 3statapack. A collection of handy Stata programs for empirical analyses
★ 59prettymaps. Draw pretty maps from OpenStreetMap data! Built with osmnx +matplotlib + shapely
★ 12kdstat. Stata module to compute summary statistics and distribution functions including standard errors and optional covariate balancing
★ 22python-edgar. Download the SEC filings index from EDGAR since 1993
★ 354here. Stata package roughly replicating the behavior of the R library "here"
★ 22pandoc-markdown-css-theme. CSS files and a template for using Pandoc to generate standalone HTML files
★ 193page_dewarp. Text page dewarping using a "cubic sheet" model
★ 1.5kscientific-visualization-book. An open access book on scientific visualization using python and matplotlib
★ 11kstata-visual-library. Inspiration and code for data visualizatio in Stata, created and maintained by DIME Analytics.
★ 96trafilatura. Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as CSV, JSON, HTML, MD, TXT, XML
★ 6.4kdid_imputation. Event studies: robust and efficient estimation, testing, and plotting
★ 202applied-methods-phd. Repo for Yale Applied Empirical Methods PHD Course
★ 2.2knotation. Collection of quotes on notation design & how it affects thought.
★ 1.9klayout-parser. A Unified Toolkit for Deep Learning Based Document Image Analysis
★ 5.8krdrobust. Robust Local Polynomial Methods for RD Designs
★ 96engineer-manager. A list of engineering manager resource links.
★ 11ksumhdfe. Summary and diagnostic information for evaluating within-fixed-effect variation.
★ 42camelot. A Python library to extract tabular data from PDFs
★ 3.8kDRDID. Doubly Robust Difference-in-Differences Estimators
★ 103typesense. Open Source alternative to Algolia + Pinecone and an Easier-to-Use alternative to ElasticSearch ⚡ 🔍 ✨ Fast, typo tolerant, in-memory fuzzy Search Engine for building delightful search experiences
★ 26ktransformers. 🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
★ 163kText-Pastry. Extend the power of multiple selections in Sublime Text. Modify selections, insert numeric sequences, incremental numbers, generate uuids, date ranges, insert continuously from a word list and more.
★ 841dataclassframe. A container for dataclasses with multi-indexing and bulk operations.
★ 314dependencies. Stata command for managing required user-written commands in a project (version freeze)
★ 5stata-economics. Economics Lesson with Stata
★ 35bbplot. R package that helps create and export ggplot2 charts in the style used by the BBC News data team
★ 1.6kngraph.hde. High dimensional embedding of a graph and its layout
★ 83practical-python. Practical Python Programming (course by @dabeaz)
★ 11kppml_fe_bias. bias corrections for PPML estimation of two-way and three-way fixed effects models
★ 13panelEvent. Estimating a Panel Event Study in Stata
★ 8covid-19-peru-data. Datos de casos confirmados, negativos, defunciones y recuperados, transcritos de los tweets del MINSA (https://twitter.com/Minsa_Peru), de sus comunicados y de su "Sala Situacional...".
★ 44did. Difference in Differences with Multiple Periods, website: https://bcallaway11.github.io/did
★ 410Datos-Abiertos-COVID-19. Datos Abiertos de COVID-19 publicados en datosabiertos.gob.pe
★ 8data-covid-minsa. limpieza rapida de datos de covid de minsa
★ 7awesome-soccer-analytics. :soccer::chart_with_upwards_trend: A curated list of awesome resources related to Soccer Analytics.
★ 617pyhdfe. High dimensional fixed effect absorption with Python 3
★ 60StataLinux. Sublime Text 3 plugin that adds support for Stata (all versions) in Linux.
★ 8manubot. Python utilities for Manubot: Manuscripts, open and automated
★ 474pandocmk. Python library that simplifies running Pandoc
★ 11pandoc-include. A pandoc filter to allow file and header inclusion
★ 82Advice. Advice (from other people) on research, graduate school, publishing, presentations, etc.
★ 23fbs-tutorial. Tutorial for creating Python/Qt GUIs with fbs
★ 2kLaplacians.jl. Algorithms inspired by graph Laplacians: linear equation solvers, sparsification, clustering, optimization, etc.
★ 245graphlayouts. new layout algorithms for network visualizations in R
★ 284adaptive. :chart_with_upwards_trend: Adaptive: parallel active learning of mathematical functions
★ 1.2ktectonic. A modernized, complete, self-contained TeX/LaTeX engine, powered by XeTeX and TeXLive.
★ 5kstata-scheme-modern. Better default plots in Stata
★ 50pandocker. 🐳 A simple docker image for pandoc with filters, templates, fonts, and the latex bazaar
★ 168gph2csv. Output plotting points from Stata graphs to CSV files
★ 3glmdr. Generalized Linear Models Done Right
★ 1R_glmhdfe. Fast Estimation of GLMs with High-Dimensional Fixed Effects
★ 29paper-tips-and-tricks. Best practice and tips & tricks to write scientific papers in LaTeX, with figures generated in Python or Matlab.
★ 3.7kppmlhdfe. Poisson pseudo-likelihood regression with multiple levels of fixed effects
★ 73stata-binscatter2. Really fast binned scatterplots in Stata
★ 42binscatter. Binscatter ggplot2 extension
★ 21panwriter. Markdown editor with pandoc integration and paginated preview.
★ 1.3kcolrspace. Stata module providing a class-based color management system in Mata
★ 5the-single-plain-text-file-cv. The single plain text file Curriculum
★ 1panflute-feedstock. A conda-smithy repository for panflute.
★ 2iA-Fonts. Free variable writing fonts from iA
★ 4.1khmda-tools. Tools to make importing and analyzing mortgage application data easier. This is a public domain work of the US Government.
★ 49ieturk. Intuitive Annotation Tool for Information Extraction / Named Entity Recognition using localturk / Amazon Mechanical Turk
★ 264Fiona. Fiona reads and writes geographic data files
★ 1.2kburd. A blue-red Stata colour scheme that supports up to 11 diverging classes.
★ 16stata-parquet. Read and write parquet files from Stata
★ 27camelot. Camelot: PDF Table Extraction for Humans
★ 3.7kstata-parquet-old. Read and write Parquet files from Stata
★ 4raster-vision-examples. Examples of using Raster Vision on open datasets
★ 174stata_kernel. A Jupyter kernel for Stata. Works with Windows, macOS, and Linux.
★ 278interesting-reads. This repo contains worthhwile essays, articles and blogposts
★ 49ggdag. :arrow_lower_left: :arrow_lower_right: An R package for working with causal directed acyclic graphs (DAGs)
★ 465alpaca. An R-package for fitting glm's with high-dimensional k-way fixed effects
★ 47lfe. Source code repository for the R package lfe on CRAN.
★ 55Terminus. Bring a real terminal to Sublime Text
★ 1.5kBibLatex-Check. A python script for checking BibLatex .bib files for common referencing mistakes!
★ 185pandoctools. Profile manager of text processing pipelines: Pandoc filters, any text CLI filters. Atom+Markdown+Pandoc+Jupyter workflow, export to ipynb. Uses Stitch fork: https://github.com/kiwi0fruit/knitty
★ 52black. The uncompromising Python code formatter
★ 42ksublime-stata. Sublime Text package for Stata (improved syntax, snippets, and shortcuts)
★ 9lala. :earth_americas: Analyze and generate reports of web logs (NGINX)
★ 60snips-nlu. Snips Python library to extract meaning from text
★ 4kscantailor. C++
★ 1.8krequests-html. Pythonic HTML Parsing for Humans™
★ 14kstata_reproducible. Tools to make output from Stata reproducible (byte-identical)
★ 14RegressionTables.jl. Journal-style regression tables
★ 138scribeAPI. scribe API
★ 81geogrid. Turning geospatial polygons into regular or hexagonal grids. For other similar functionality see the tilemaps package https://github.com/kaerosen/tilemaps
★ 405osrm-backend. Open Source Routing Machine - C++ backend
★ 7.9kstarpolishr. Post-polishing of stagazer output
★ 25pandoc-mustache. Pandoc filter for variable substitution using Mustache syntax
★ 58profile-summary-for-github. Tool for visualizing GitHub profiles
★ 20kcurio. Good Curio!
★ 4.1kebook. A template project for building an eBook, using Python, Pandoc and Markdown.
★ 77netgraph. Publication-quality network visualisations in python
★ 747DeepLearningImplementations. Implementation of recent Deep Learning papers
★ 1.8kscatteract. Project which implements extraction of data from scatter plots
★ 216gcv2hocr. gcv2hocr converts from Google Cloud Vision OCR output to hocr to make a searchable pdf.
★ 108siunitx. A comprehensive (SI) units package for LaTeX
★ 403tup. Tup is a file-based build system.
★ 1.3kcalipmatch. Stata module for caliper matching without replacement
★ 2datasette. An open source multi-tool for exploring and publishing data
★ 11kStataRegex. A Stata implementation of the Java regular expression utilities
★ 3ts2sls. Two-sample two-stage least squares estimation
★ 11charts. Simple, responsive, modern SVG Charts with zero dependencies
★ 15kgantt. Open Source Javascript Gantt
★ 6.1klanguage-stata. Syntax highlighting for Stata in Atom
★ 50engrafo. Convert LaTeX documents into beautiful responsive web pages using LaTeXML.
★ 1.1kgraphs-in-machine-learning. A curated list of resources on the intersection of Graphs and Machine Learning
★ 30neva. Network valuation in financial systems
★ 39rcall. Seamless interactive R in Stata. rcall allows communicating data sets, matrices, variables, and scalars between Stata and R conveniently
★ 98elasticregress. Stata implementation of the Friedman, Hastie and Tibshirani (2010, JStatSoft) coordinate descent algorithm for elastic net regression
★ 15linemap. Create maps made of (ridge) lines
★ 114GLM.jl. Generalized linear models in Julia
★ 636wand. The ctypes-based simple ImageMagick binding for Python
★ 1.5kbinscatter. Stata module to generate binned scatterplots —
★ 45Historical-Populations. Historical US City populations
★ 43fastLink. R package fastLink: Fast Probabilistic Record Linkage
★ 293USAboundaries. Historical and Contemporary Boundaries of the United States of America
★ 68synth_runner. A tool to run a pool of synthetic controls, conduct inference, and produce visualizations.
★ 45boottest. Stata module for fast wild bootstrap-based inference. Releases posted here are appropriate for use, and are usually posted promptly on SSC.
★ 13stata-gtools. Faster implementation of Stata's collapse, reshape, xtile, egen, isid, and more using C plugins
★ 195tidy. An implementation of the tidyr package in Stata
★ 14snappydata. Project SnappyData - memory optimized analytics database, based on Apache Spark™ and Apache Geode™. Stream, Transact, Analyze, Predict in one cluster
★ 1kline_profiler. (OLD REPO) Line-by-line profiling for Python - Current repo ->
★ 3.6kprofilehooks. Python decorators for profiling/tracing/timing a single function
★ 338football-crunching. Analysis and datasets about football (soccer)
★ 329historical-us-city-populations. Historical city populations
★ 33PySolFC. A comprehensive, feature-rich, open source, and portable, collection of Solitaire games.
★ 557useful-stata. Do file that quickly installs useful Stata programs
★ 7Beamer-Theme-Execushares. A minimalist and modern Beamer theme
★ 298map-vectorizer. An open-source map vectorizer
★ 659haskell-style-guide. A style guide for Haskell code.
★ 958pdftabextract. A set of tools for extracting tables from PDF files helping to do data mining on (OCR-processed) scanned documents.
★ 2.3kcerebro. 🔵 Cerebro is an open-source launcher to improve your productivity and efficiency
★ 8.6kpdfminer.six. Community maintained fork of pdfminer - we fathom PDF
★ 7kpdfminer. Python PDF Parser (Not actively maintained). Check out pdfminer.six.
★ 5.3kcmark. CommonMark parsing and rendering library and program in C
★ 2kdata-science-ipython-notebooks. Data science Python notebooks: Deep learning (TensorFlow, Theano, Caffe, Keras), scikit-learn, Kaggle, big data (Spark, Hadoop MapReduce, HDFS), matplotlib, pandas, NumPy, SciPy, Python essentials, AWS, and various command lines.
★ 29kNakedTensor. Bare bone examples of machine learning in TensorFlow
★ 2.4kbcrpuse. Stata module to Import data from the Peruvian Central Bank (BCRP)
★ 2PandasDataFrameGUI. A minimalistic GUI for analyzing Pandas DataFrames.
★ 326boilr. :zap: boilerplate template manager that generates files or directories from template repositories
★ 1.8kpandocomatic. Automate the use of pandoc
★ 173ivreghdfe. Run IV/2SLS with many levels of fixed effects (i.e. ivreg2+reghdfe)
★ 112ggalt. :earth_americas: Extra Coordinate Systems, Geoms, Statistical Transformations & Scales for 'ggplot2'
★ 690hrbrthemes. :lock_with_ink_pen: Opinionated, typographic-centric ggplot2 themes and theme components
★ 1.4kreverse-geocoder. A fast, offline reverse geocoder in Python
★ 1.9kcpython. The Python programming language
★ 74kstatsmodels. Statsmodels: statistical modeling and econometrics in Python
★ 12kGeospatial_Data_with_Python. Introduction to Geospatial Data with Python
★ 182jupyter-themes. Custom Jupyter Notebook Themes
★ 9.8klegit. Git for Humans, Inspired by GitHub for Mac™.
★ 5.7kalpine-python. Python images for amd64, arm32v6 and arm32v7 based on Alpine Linux (3.6, 2.7)
★ 39latexrun. A 21st century LaTeX wrapper
★ 637pandocpm. Manage the install/update/uninstall of packages
★ 7rinohtype. The Python document processor
★ 529pypandoc. Thin wrapper for "pandoc" (MIT)
★ 1.1kpantable. CSV Tables in Markdown: Pandoc Filter for CSV Tables
★ 93github. a module for building, searching, installing, managing, and mining Stata packages from GitHub
★ 111paru. Control pandoc with Ruby and write pandoc filters in Ruby
★ 41Folder-Structure-Conventions. Folder / directory structure options and naming conventions for software projects
★ 2kpandoc-filter-test. Playing with pandoc filters to translate unicode literals
★ 2tint. Tint is not Tufte
★ 279statastan. Stata interface for Stan.
★ 22ceres-solver. A large scale non-linear optimization library
★ 4.5ktypora-issues. Bugs, suggestions or free discussions about the minimal markdown editor — Typora
★ 1.6kjanitor. simple tools for data cleaning in R
★ 1.5kforcats. 🐈🐈🐈🐈: tools for working with categorical variables (factors)
★ 558xsv. A fast CSV command line toolkit written in Rust.
★ 11kwhen-changed. Execute a command when a file is changed
★ 1.2krclone. "rsync for cloud storage" - Google Drive, S3, Dropbox, Backblaze B2, One Drive, Swift, Hubic, Wasabi, Google Cloud Storage, Azure Blob, Azure Files, Yandex Files
★ 59k