Skip to content

Adding nnlib integration#2

Merged
dijopaul merged 4 commits intomainfrom
dijopaul_add-nnlib2
Jul 26, 2024
Merged

Adding nnlib integration#2
dijopaul merged 4 commits intomainfrom
dijopaul_add-nnlib2

Conversation

@dijopaul
Copy link
Owner

No description provided.

@dijopaul dijopaul merged commit 0f8c702 into main Jul 26, 2024
dijopaul pushed a commit that referenced this pull request Sep 19, 2024
Summary:
Fix remaining warnings when building MPS delegate with older iOS SDKs

cc shoumikhin, cccclai

Pull Request resolved: pytorch#5365

Reviewed By: kirklandsign

Differential Revision: D62668033

Pulled By: shoumikhin

fbshipit-source-id: efcb8449b1f38259914c76337e12dc38487bf36b
@dijopaul dijopaul deleted the dijopaul_add-nnlib2 branch November 28, 2024 09:52
dijopaul pushed a commit that referenced this pull request Jun 27, 2025
Differential Revision: D75034439

Pull Request resolved: pytorch#11011
dijopaul pushed a commit that referenced this pull request Jun 27, 2025
Differential Revision: D75982351

Pull Request resolved: pytorch#11456
dijopaul pushed a commit that referenced this pull request Oct 20, 2025
BNNS copy crashes the process when the dtypes differ
(pytorch#11714).

With the example in this PR
(pytorch#11714), we crash the
process on main. Here is the stack trace from LLDB:

```
Process 19234 stopped
* thread #1, queue = 'com.apple.main-thread', stop reason = signal SIGABRT
    frame #0: 0x0000000190ac9388 libsystem_kernel.dylib`__pthread_kill + 8
libsystem_kernel.dylib`__pthread_kill:
->  0x190ac9388 <+8>:  b.lo   0x190ac93a8    ; <+40>
    0x190ac938c <+12>: pacibsp 
    0x190ac9390 <+16>: stp    x29, x30, [sp, #-0x10]!
    0x190ac9394 <+20>: mov    x29, sp
(lldb) bt
* thread #1, queue = 'com.apple.main-thread', stop reason = signal SIGABRT
  * frame #0: 0x0000000190ac9388 libsystem_kernel.dylib`__pthread_kill + 8
    frame #1: 0x0000000190b0288c libsystem_pthread.dylib`pthread_kill + 296
    frame #2: 0x0000000190a0bc60 libsystem_c.dylib`abort + 124
    frame #3: 0x0000000190910174 libsystem_malloc.dylib`malloc_vreport + 892
    frame #4: 0x0000000190913c90 libsystem_malloc.dylib`malloc_report + 64
    frame #5: 0x000000019091821c libsystem_malloc.dylib`___BUG_IN_CLIENT_OF_LIBMALLOC_POINTER_BEING_FREED_WAS_NOT_ALLOCATED + 32
    frame #6: 0x000000019d2f4084 libBNNS.dylib`___lldb_unnamed_symbol1620 + 564
    frame #7: 0x000000019d2f5bac libBNNS.dylib`___lldb_unnamed_symbol1628 + 680
    frame #8: 0x000000019d69ce48 libBNNS.dylib`BNNSCopy + 616
    frame #9: 0x000000030c74d950 _portable_lib.cpython-310-darwin.so`(anonymous namespace)::copy_using_bnns(executorchcoreml::MultiArray const&, executorchcoreml::MultiArray&) + 188
    frame #10: 0x000000030c74cfdc _portable_lib.cpython-310-darwin.so`(anonymous namespace)::copy(executorchcoreml::MultiArray const&, executorchcoreml::MultiArray&, executorchcoreml::MultiArray::CopyOptions) + 72
    frame #11: 0x000000030c74ceec _portable_lib.cpython-310-darwin.so`executorchcoreml::MultiArray::copy(executorchcoreml::MultiArray&, executorchcoreml::MultiArray::CopyOptions) const + 148
    frame #12: 0x000000030c7488d4 _portable_lib.cpython-310-darwin.so`invocation function for block in (anonymous namespace)::copy(MLMultiArray*, executorchcoreml::MultiArray&) + 376
    frame #13: 0x000000030c748ac8 _portable_lib.cpython-310-darwin.so`invocation function for block in (anonymous namespace)::copy(MLMultiArray*, executorchcoreml::MultiArray&) + 52
    frame #14: 0x000000019ad33f4c CoreML`CoreML::MultiArrayBuffer::getBytesWithHandler(void (void const*, unsigned long) block_pointer) const + 340
    frame #15: 0x000000019ad34138 CoreML`-[MLMultiArray(ScopedBufferAccess) getBytesWithHandler:] + 152
    frame #16: 0x000000030c7485ec _portable_lib.cpython-310-darwin.so`(anonymous namespace)::copy(MLMultiArray*, executorchcoreml::MultiArray&) + 296
    frame #17: 0x000000030c744f68 _portable_lib.cpython-310-darwin.so`(anonymous namespace)::set_outputs(std::__1::vector<executorchcoreml::MultiArray, std::__1::allocator<executorchcoreml::MultiArray>>&, NSArray<MLMultiArray*>*) + 180
```


With this PR, the process succeeds.
dijopaul pushed a commit that referenced this pull request Oct 20, 2025
### Summary
- pytorch#14048 > add quantized test case with GLU decomposition
- pytorch#14049 > add e2e example where constant expansion is applied
- pytorch#14050 > add e2e example and source transform for 6D operation
- pytorch#14051 > add e2e example and complement missed annotation
- pytorch#14052 > add e2e example and dedicated passe for 6D partition


Fixes pytorch#14048
Fixes pytorch#14049
Fixes pytorch#14050
Fixes pytorch#14051
Fixes pytorch#14052

### Test plan
MATRIX = {convnext_small, maxvit_t, swin_v2_t, vit_b_16}
```bash
python backends/qualcomm/tests/test_qnn_delegate.py TestExampleOssScript.test_${MATRIX} -b build-android/ -m SM8750 -s $SN -a /path/to/test_artifacts/ -i /path/to/imagenet_1k/imagenet-mini/val -r .
```
```bash
python backends/qualcomm/tests/test_qnn_delegate.py TestQuantizedModel.test_qnn_backend_conformer -b build-android/ -m SM8750 -s $SN -a /path/to/test_artifacts/
```
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant