Commits · 28ee806dcbd803f4079365dd308a673bd1a89588 · wenyuanbo / tic

05 Feb, 2020 1 commit
- allow customize mkldnn library location (#4814) · 3e7bd703
  Haichen Shen committed 5 years ago
  
  3e7bd703 Browse Directory
27 Jan, 2020 1 commit
- [Build] Explicitly link to cublasLt if it exists (#4776) · 00ec7f9c
```
* Explicitly link to cublasLt

* Only link cublasLt if it's found

Co-authored-by: Jon Soifer <jonso@microsoft.com>
```
  Jon Soifer committed 5 years ago
  00ec7f9c Browse Directory
19 Jan, 2020 1 commit

[REFACTOR][CODEGEN] codegen->target, build_module->driver (#4742) · 33b0831c

This PR moves the codegen related code into the target folder,
as they are target specific functionalities.

We also adopt the term "compiler driver" in common compiler infra
such as rust, GHC and clang.
As a result, build_module is moved into the driver folder.

committed 5 years ago

33b0831c Browse Directory

16 Jan, 2020 1 commit
- [Runtime] EdgeTPU runtime for Coral Boards (#4698) · 31021d2b
  Thierry Moreau committed 5 years ago
  
  31021d2b Browse Directory
10 Jan, 2020 1 commit
- [CodeGen] Generate blob use LLVM directly (#4657) · 54fb55a1
  Zhao Wu committed 5 years ago
  
  54fb55a1 Browse Directory
18 Dec, 2019 1 commit
- [Relay] External codegen (#4482) · c44b7bf1
  Zhi committed 5 years ago
  
  c44b7bf1 Browse Directory
04 Dec, 2019 1 commit
- [CONTRIB] TFLite Runtime (#4439) · 24713bde
  ziheng committed 5 years ago
  
  24713bde Browse Directory
02 Dec, 2019 1 commit
- a tiny typo (#4452) · 11af82c0
  HarryWu committed 5 years ago
  
  11af82c0 Browse Directory
22 Nov, 2019 1 commit
- [TVM][RUNTIME] A minimum example to generate external library wrappers for DSOModule (#4280) · e0810512
  Zhi committed 5 years ago
  
  e0810512 Browse Directory
18 Nov, 2019 1 commit
- [SOURCE] Add ASF header to __init__.py files (#4359) · 00521fab
  Tianqi Chen committed 5 years ago
  
  00521fab Browse Directory
15 Nov, 2019 1 commit
- [Contrib] Add MKL DNN option (#4323) · 72821b20
```
* [Contrib] Add MKL DNN

* update

* update
```
  Haichen Shen committed 5 years ago
  72821b20 Browse Directory
31 Oct, 2019 2 commits
- [BUILD] Disable utvm standalone runtime by default (#4240) · a3ca1a4d
  Tianqi Chen committed 5 years ago
  
  a3ca1a4d Browse Directory
- [CI] Update GPU docker to cuda10 (#4228) · b2155f70
```
* [CI] Update the ci-gpu to use cuda10

* [CI] Enforce tensorcore gpu for unittest
```
  Tianqi Chen committed 5 years ago
  b2155f70 Browse Directory
30 Oct, 2019 1 commit
- [CI] use llvm9 for the gpu tests (#4224) · 83385d42
```
* [CI] use llvm9 for the gpu tests

* Update Docker script to support new nvidia docker
```
  Tianqi Chen committed 5 years ago
  83385d42 Browse Directory
27 Oct, 2019 1 commit
- [RUNTIME] Separate runtime related contrib into runtime/contrib (#4207) · dcc6af53
  Tianqi Chen committed 5 years ago
  
  dcc6af53 Browse Directory
24 Oct, 2019 1 commit
- [cmake][ANTLR] Support setting path to ANTLR jar (#4176) · 2460f908
```
* Support setting path to ANTLR jar

* Update comment
```
  Jon Soifer committed 5 years ago
  2460f908 Browse Directory
20 Oct, 2019 1 commit
- [Runtime] Enable option to use OpenMP thread pool (#4089) · 97ea31c8
  Haichen Shen committed 5 years ago
  
  97ea31c8 Browse Directory
08 Oct, 2019 1 commit

[VTA] hotfix for de10-nano driver (#4081) · be869546

Issue:
git clone latest TVM/VTA and run VTA on xilinx FPGA board, application
crashed due to the "call stack overflow" which caused by a infinite recursive
function call. this issue ever happen before and get addressed by PR 3843.

Analysis:
seems like de10-nano driver PR  used old code base then the  logic change
of 3843 get eliminated.

Solution:
add the logic back.

committed 5 years ago

be869546 Browse Directory

17 Sep, 2019 1 commit
- More friendly error msg; Fix Android Demo LLVM ver (#3962) · a3073457
  Junru Shao committed 5 years ago
  
  a3073457 Browse Directory
14 Sep, 2019 1 commit
- trivial (#3954) · e35e1cc2
  Junru Shao committed 5 years ago
  
  e35e1cc2 Browse Directory
13 Sep, 2019 1 commit
- Vulkan2 Runtime API (#3849) · 2536465c
  Andrew Tulloch committed 5 years ago
  
  2536465c Browse Directory
12 Sep, 2019 1 commit

[RFC] [Contrib] Minimal runtime (~12kb .text on ARMv7/x86) for subset of TVM models (#3567) · 1de52bb0

This is an alternative implementation of a subset of the TVM runtime API (and
graph runtime) that focuses entirely on reducing code size, at the expense of
functionality (no tvm.extern(..) calls via PackedFunc, CPU only, etc). It might
be worth incrementally expanding the surface area if there's interest.

The motivation for this work was seeing what the minimal useful subset of the
TVM runtime is. This is relevant for e.g. super code-size constrained
applications in e.g. embedded/mobile. The current runtime is more like O(100KiB)
or so, so this might be compelling for some users.

The smaller surface area for auditing might make this relevant for
https://github.com/dmlc/tvm/issues/3159, or the usecases I was thinking about in
https://github.com/dmlc/tvm/issues/2523#issuecomment-459165815 re: the Rust
runtime.

The symbols in the tvm::minimalruntime space (i.e. excluding std:: and
picojson::) are about 5KiB, so I think there's a bunch of room here (i.e. we
could replace picojson:: with [`jsmn`](https://zserge.com/jsmn.html) or
something, and we could replace more of the `std::unordered_map` usage, etc with
custom primitives as well (similar to the `DynArray`).

committed 5 years ago

1de52bb0 Browse Directory

07 Sep, 2019 2 commits

[Fix] Fix blas cmake for mac os (#3898) · 7a15aedf
```
* fix cmake for mac os

* rename
```
Haichen Shen committed 5 years ago
7a15aedf Browse Directory

[VTA] Support TLPP in function simulator. (#3555) · 50c4546f

* [VTA] Support TLPP in function simulator.
Issue:
currently vta function simulator just doing serialized instruction
execution, the dependency logic of runtime ISA which use for task
level pipe line parallelism can not get verified by function simulator.

Solution:
make the simulator driver to be multiple thread and support TLPP.

Benefit:
TLPP support VTA function simulator would make VTA logic testing/debug
/change more easy.

replace boost lockfree queue

add configure control for simulator tlpp enable or disable.

change code tyle into google style.

Wrap queue read/write and sync logic to make function call more simple.

Add some comments.

Remove MT logic, change into Single thread mode.

address review comments.

code style change to match google code style and add comments.

add cmake macro to enable/disable simulator tlpp logic.

submodule update.

correct file name mentioned in comments.

* remove USE_VTA_FSIM_TLPP.

committed 5 years ago

50c4546f Browse Directory

06 Sep, 2019 1 commit
- Add another MKL name alias for MKL (#3853) · d464d2ba
```
Installed through pypi
```
  Jason Knight committed 5 years ago
  d464d2ba Browse Directory
05 Sep, 2019 1 commit

[VTA] de10-nano driver (#3394) · 734df8d5

* rework;

* `de10-nano` -> `de10nano`;

* fix compilation error;

* bug fix;

* Update install.md

* Update install.md

* Update install.md

* update with current runtime;

* add debug messages;

* bug fix in cma kernel module;

committed 5 years ago

734df8d5 Browse Directory

02 Sep, 2019 1 commit
- [VTA][Chisel] rename USE_TSIM macro with USE_VTA64 and cleanup runtime (#3872) · 4434a89c
  Luis Vega committed 5 years ago
  
  4434a89c Browse Directory
29 Aug, 2019 2 commits

[VTA] Infinite recursive device_api.ext_dev call fix. (#3843) · 61d19ccc

Issue
when try vta on fpga board, would see a Infinite recursive
device_api.ext_dev issue that cause stack overflow and vta
failed.

Analysis:
device_api.ext_dev function in rpc_server.py is use to load
vta library, once vta library get load, device_api.ext_dev would
get replaced with vta function by vta library, vta device_api.cc
did such work, but because a logic issue in VTA.cmake, the said file
not get compiled, then vta would keep failing on rpc_server.py.

Solution:
fix the logic issue in VTA.cmake.

committed 5 years ago

61d19ccc Browse Directory

Support MKL on Windows (#3837) · 710ac146
Jon Soifer committed 5 years ago

710ac146 Browse Directory

26 Aug, 2019 1 commit

[VTA][TSIM] Introduce Virtual Memory for TSIM Driver (#3686) · 92b6ca71

* initial virtual memory;

* initial integration;

* include the header file in cmake;

* implement allocation with virtual to logical address mapping;

* virtual memory for tsim_driver;

* implement the missing memory release function;

* readability improvement;

* readability improvement;

* address review comments;

* improved robustness in virtual memory allocation;

* remove VTA_TSIM_USE_VIRTUAL_MEMORY macro and use virtual memory for tsim by default;

* link tvm against vta library;

* merge with master

* build virtual memory system without linking tvm against vta;

* minor change;

* reuse VTA_PAGE_BYTES;

* using DRAM class from sim_driver as VirtualMemoryManager;

* satisfy linter;

* add comments in code;

* undo changes to Makefile

* undo changes to Makefile

* retrigger ci;

* retrigger ci;

* directly call into VirtualMemoryManager::Global()

committed 5 years ago

92b6ca71 Browse Directory

21 Aug, 2019 1 commit

[Relay][VM]VM Profiler (#3727) · 95f12e31

* [Relay][VM]VM debugger

* Report mean/min/max for op duration

* Typos

* Lint

* Lint

* Lint

* Support build debug VM in CMake

* Lint

* Enable VM debug in unit test

* Disable debug vm test until new docker image is built

* Add device sync code

* Fix qnn unit test

* Disable vm debug by default

* Rename files

* Rename classes

* Fix comment

* Fix comment

committed 5 years ago

95f12e31 Browse Directory

13 Aug, 2019 1 commit

[VTA][TSIM][Build] Towards TSIM CI testing (#3704) · e518fe1c

* building TSIM specific library along with fast simulator to quickly switch between dlls

* cmake controlled TSIM libraries

* always build tsim driver in either simulation modes

* build DLLs based on CMAKE flags

* updating the jenkinsfile

* small restructuring

* reducing the cmake flags

* update instructions

* reverting to 3 flags

* update Jenkinsfile

* adding new line

* enabling TSIM unit and integration tests

* fix description

* temporarily disabling task_python_vta tests in CPU Build stage

* move CPU tests in unit test stage

* stage  reorg

* better make

* disabling TSIM tests for now

* reverting some restructuring

* fix

committed 5 years ago

e518fe1c Browse Directory

29 Jul, 2019 2 commits

[VTA] [CMake] hotfix tsim rules (#3650) · 6970fc30
Luis Vega committed 5 years ago

6970fc30 Browse Directory

[VTA] Refactor to increase platform coverage (Ultra96 etc.) (#3496) · f55609b4

* hardware refactor for increased FPGA coverage, small optimizations

* fix header

* cleaning up parameters that won't be needed for now

* streamlining makefile, and simplifying tcl scripts

* moving parameter derivation into pkg_config.py, keeping tcl scripts lightweight

* refactoring tcl script to avoid global variables

* deriving AXI signals in pkg_config.py

* unifying address map definition for hardware and software drivers

* single channel design for ultra96 to simplify build

* enable alu by default, no mul opcode for now

* hardware fix

* new bitstream; vta version

* avoid error when env variable is not set

* ultra96 cleanup

* further cleaning up tcl script for bitstream generation

* preliminary rpc server support on ultra96

* rpc server tracker scripts

* ultra96 ldflag

* ultra96 support

* ultra96 support

* cleanup line

* cmake support for ultra96

* simplify memory instantiation

* cleaning up IP parameter initialization

* fix queue instantiation

* 2019.1 transition

* fix macro def

* removing bus width from config

* cleanup

* fix

* turning off testing for now

* cleanup ultra96 ps insantiation

* minor refactor

* adding comments

* upgrading to tophub v0.6

* model used in TVM target now refers to a specific version of VTA for better autoTVM scheduling

* revert change due to bug

* rename driver files to be for zynq-type devices

* streamlining address mapping

* unifying register map offset values between driver and hardware generator

* rely on cma library for cache flush/invalidation

* coherence management

* not make buffer packing depend on data types that can be wider than 64bits

* refactor config derivation to minimize free parameters

* fix environment/pkg config interaction

* adding cfg dump property to pkgconfig:

* fix rpc reconfig

* fix spacing

* cleanup

* fix spacing

* long line fix

* fix spacing and lint

* fix line length

* cmake fix

* environment fix

* renaming after pynq since the driver stack relies on the pynq library - see pynq.io

* update doc

* adding parameterization to  name

* space

* removing reg width

* vta RPC

* update doc on how to edit vta_config.json

* fix path

* fix path

committed 5 years ago

f55609b4 Browse Directory

26 Jul, 2019 1 commit
- Make Google Test usage configurable in CMake files (#3628) · f5464ce2
```
* Add USE_GTEST as a CMake variable

* Add GTest section in installation docs

* Incorporate feedback
```
  Logan Weber committed 5 years ago
  f5464ce2 Browse Directory
25 Jul, 2019 1 commit

Implementation of uTVM (#3227) · ef909df1

* uTVM interfaces (#14)

* some minor interface changes

* implemented HostLowLevelDevice

* added MicroDeviceAPI

* implemented micro_common and added Python interfaces

* current status, semi implemented micro session

* added micro_common implementation and python interfaces (#18)

* added micro_common implementation and python interfaces (#18)

* current status, semi implemented

* host test working

* updated interfaces for MicroSession arguments allocation

* make somewhat lint compatible

* fix based on comments

* added rounding macro

* fix minor bug

* improvements based on comments

* Clean up `binutil.py` and make Python-3-compatible

* Change argument allocation design

* Address feedback and lint errors

* Improve binutil tests

* Simplify allocator (per @tqchen's suggestions)

* Doc/style fixes

* farts

* mcgee

* rodata section werks

(and so does `test_runtime_micro_workspace.py`)

* simple graph runtime werk

* TEMP

* ResNet works, yo

* First round of cleanup

* More cleanup

* runs a dyson over the code

* Another pass

* Fix `make lint` issues

* ready to pr... probably

* final

* Undo change

* Fix rebase resolution

* Minor fixes

* Undo changes to C codegen tests

* Add `obj_path` in `create_micro_lib`

* TEMP

* Address feedback

* Add missing TODO

* Partially address feedback

* Fix headers

* Switch to enum class for `SectionKind`

* Add missing ASF header

* Fix lint

* Fix lint again

* Fix lint

* Kill lint warnings

* Address feedback

* Change Python interface to MicroTVM

All interaction with the device is now through `Session` objects, which
are used through Python's `with` blocks.

* Reorder LowLevelDevice interface

* Store shared ptr to session in all alloced objects

* Move helper functions out of `tvm.micro`

* Switch static char arr to vector

* Improve general infra and code quality

Does not yet address all of tqchen's feedback

* Forgot a rename

* Fix lint

* Add ASF header

* Fix lint

* Partially address MarisaKirisame's feedback

* Lint

* Expose `MicroSession` as a node to Python

* Revert to using `Session` constructor

* Fix compiler error

* (Maybe) fix CI error

* Debugging

* Remove

* Quell lint

* Switch to stack-based session contexts

* Make uTVM less intrusive to host codegen

And use SSA for operands of generated ternary operators

* Inline UTVMArgs into UTVMTask struct

* Remove `HostLowLevelDevice` header

* Remove `BaseAddr` class

* Address feedback

* Add "utvm" prefix to global vars in runtime

* Fix lint

* Fix CI

* Fix `test_binutil.py`

* Fix submodules

* Remove ResNet tests

* Make `test_binutil.py` work with nose

* Fix CI

* I swear this actually fixes the binutil tests

* lint

* lint

* Add fcompile-compatible cross-compile func

* Add docs for uTVM runtime files

* Move pointer patching into `MicroSession`

* Fix lint

* First attempt at unifying cross-compile APIs

* Fix lint

* Rename `cross_compile` back to `cc`

* Address feedback

* Remove commented code

* Lint

* Figure out failing function

* Remove debugging code

* Change "micro_dev" target to "micro"

* Add checks in tests for whether uTVM is enabled

* Add TODO for 32-bit support

* Rename more "micro_dev" to "micro"

* Undo rename

We already have `tvm.micro` as a namespace.  Can't have it as a method
as well.

* Fix failing CI

Thanks to @tqchen for finding this bug.  Emitting ternary operators for
`min` and `max` causes concurrency bugs in CUDA, so we're moving the
ternary op emissions from `CodeGenC` to `CodeGenCHost`.

* Address feedback

* Fix lint

committed 5 years ago

ef909df1 Browse Directory

09 Jul, 2019 1 commit
- [Relay] Roundtrip part of pretty printer and parser (#3460) · 7fb9557b
```
* init

fix rebase

lint

fix cmake

try again

fix ci

* add gitignore

* fix format

* do not include .interp and .tokens
```
  雾雨魔理沙 committed 5 years ago
  7fb9557b Browse Directory
28 Jun, 2019 1 commit
- [Relay][Parser] simplify build script, remove python 2 support (#3419) · b05dc12f
```
* simplify build script, remove python 2 support

* remove py2 file

* update py3
```
  雾雨魔理沙 committed 5 years ago
  b05dc12f Browse Directory
05 Jun, 2019 1 commit
- [VTA] [Hardware] Chisel implementation (#3258) · 32f74f31
  Luis Vega committed 5 years ago
  
  32f74f31 Browse Directory
01 Jun, 2019 1 commit

[Bugfix][VTA] PkgConfig cause crash in PYNQ board due to link library (#3257) · f6acf2e5

* [Bugfix][VTA] PkgConfig cause crash in PYNQ board due to link library
not exist.

Symptom:
When run vta_get_started.py with pynq board, host crash and
complain "cannot find -lsds_lib" and "cannot find -l:libdma.so"

Reproduce:
At pynq board, delete the ./build/vta_config.json, then run rpc
server.
In host machine run vta_get_started.py, issue would reproduce.

Analysis:
This issue caused by 'PkgConfig' function  still using pynq2.1
library which not exist in pynq2.4 anymore, when a "reconfig_runtime"
logic of rpc_server.py get triggered , the compile would failed due to
link library not exist.

Solution:
change the link library to libcma.so.

* [Document Change][VTA] Change pynq version from 2.3 into 2.4.

Issue:
pynq 2.3 image not available anymore from pynq download page and pynq
2.4 is the current latest image which available in the said website, after
verification, currently VTA work good with pynq 2.4 image, hence update
related document from pynq 2.3 to 2.4.

committed 5 years ago

f6acf2e5 Browse Directory